Qwen3.8-LiveTranslate: Names the speaker. Carries the meaning.

The model reduces average lagging to 2.3 seconds across 60 input languages while adding real-time multi-speaker separation and voice cloning.

Developers building live interpretation tools can stream multi-speaker speech translations with voice cloning without managing separate ASR and diarization pipelines.

Sources

Read this as text

Back to the AI news