NVIDIA Expands AI for Media with Frame Generation and Lip Sync Tools
NVIDIA is building out a single toolset for broadcasters that covers capture, motion, upscaling, and audio, which points to production pipelines where real-time translation and dubbing run alongside the video processing itself.
Reporting from 1 source: GIGAZINE.
NVIDIA added several AI tools to its AI for Media collection for the broadcast industry in September 2026. The new additions include 3D Body Pose, which converts joint positions into 2D and 3D motion data, Video Frame Generation for raising frame rates two to four times, Video Super Resolution for upscaling, TrueHDR for real-time SDR to HDR conversion, LipSync for matching mouth movements to audio, Active Speaker Detection, and Studio Voice for reducing noise and reverberation.
NVIDIA 3D Body Pose estimates joint positions and angles and turns them into structured 2D and 3D data without markers on the subject, which the source suggests for tracking athletes during replays and for animation production. Video Frame Generation inserts generated frames between existing ones to raise frame rates by two to four times, and work continues toward 8x slow motion. Video Super Resolution upscales while reducing noise, blur, and compression artifacts, with a real-time mode and a quality mode. TrueHDR converts SDR to HDR in real time at up to about 2000 nits.
- NVIDIA 3D Body Pose: Estimates joint positions and angles as structured 2D and 3D data without markers
- Video Frame Generation (VFG): Inserts generated frames to raise frame rates two to four times
- NVIDIA Video Super Resolution (VSR): Upscales video and reduces noise, blur, and compression artifacts
- NVIDIA TrueHDR: Converts SDR to HDR in real time at up to about 2000 nits
- NVIDIA LipSync: Changes mouth movements to match input audio
- Active Speaker Detection: Identifies multiple speakers in an audio file
- Studio Voice: Reduces noise and reverberation in recorded data
Synthesized by Yomimono from the 1 cited source below, including Japanese-language reporting where cited, then editorially reviewed before publishing.