Inference Performance Optimization
How sync. uses Modal to lipsync 100 hours of video a day

How sync. uses Modal to lipsync 100 hours of video a day

4/18/2025

What this post added

This post details how sync. uses Modal's parallel execution model to process over 100 hours of video per day for lipsyncing. The workflow involves splitting videos into scenes, running face detection and translation on T4 GPU containers, processing scenes with proprietary LipSync GenAI models on A100 GPUs, and then stitching the dubbed scenes back together. This demonstrates the platform's capability for large-scale, compute-intensive video processing workloads.

Read the original post ↗