Real-Time Video Translator for Captioned Videos
A real-time video translator starts producing translated speech as playback advances. That is different from Standard mode, which generates a synchronized track before playing it. It is also different from creator tools that render and export a new lip-synced video.
In Unlimited Universal Video Translator, real-time mode is a PRO feature. The free path uses Standard mode, Google Translate, and an on-device voice. Both modes need a supported player with detectable captions; neither one transcribes raw audio.
Standard mode versus real-time mode
Standard mode prepares translated speech from the caption track and builds a synchronized audio track. The initial wait depends on video length, number of caption segments, selected voice, and the speed of the device. Once prepared, the track follows play, pause, and seek events.
Real-time mode uses a compatible voice engine to generate speech closer to playback. It reduces the up-front wait, but it still has to:
- detect the existing captions;
- translate upcoming segments;
- generate speech with the selected engine;
- keep that speech aligned with the player.
“Real-time” therefore means a lower-latency viewing pipeline, not simultaneous interpretation of arbitrary live audio.
What it does not do
- It does not create captions for a video that has none.
- It does not promise dubbing for live streams.
- It does not bypass DRM or unsupported players.
- It does not alter the speaker’s lips.
- It does not export a translated video file.
If you need a finished asset for publishing, choose an upload-and-render localization tool. If you need to understand a supported course or tutorial while watching it, a browser dubber is the closer fit.
How to try real-time mode
- Install Unlimited Universal Video Translator.
- Open a supported video and enable its captions.
- Confirm that the floating Dub control detects the video and subtitle track.
- Open the mode and voice settings.
- Select a realtime-capable PRO engine and your target language.
- Start playback and give the first segment time to prepare.
If the Dub control does not appear, turn captions on and reload the page. A player may contain visible subtitles without exposing them in a form the extension can currently detect.
Privacy depends on the selected engines
The extension does not upload the video file. Processing is not automatically fully local, though:
- local speech generation stays on the device;
- Google Translate receives the caption text needed for translation;
- cloud TTS receives the text it needs to speak;
- no-key LLM and BYOK providers receive the text required by that provider.
Choose Standard mode with a local voice when local speech generation matters more than start-up speed. Choose real-time mode when lower latency is worth using a PRO-compatible engine.
For the free workflow, see Free Video Translator. For the complete caption-first process, read How to Translate a Video.
Related guides
The AI dubbing pipeline overview explains what this mode sits on top of, and the Chrome extension guide covers getting it installed. If real-time turns out not to be what you need, the output-first decision framework is worth a read.
Frequently asked questions
Is real-time video translation free?
Not in this product. The free path uses Standard mode. Real-time mode is included with PRO-supported engines.
Can it translate a live video call or livestream?
It is not promised for that use case. The extension needs detectable captions on a supported player, and live streams may not expose a usable caption track.
Does it work without subtitles?
No. There is no built-in speech-to-text or Whisper pipeline in the current release.
Does real-time mean zero delay?
No. Translation and speech generation still take time. The mode reduces the up-front wait by processing closer to playback.
Can I download the result?
No. The translated subtitles and speech are part of the browser viewing experience, not a rendered output file.
