TREE NEWS update: Qwen released Qwen3.8-LiveTranslate, a new simultaneous interpretation large model built on an Interleave architecture, with accuracy, fluency and conciseness all improved. Average latency per character (LAAL) fell from 2.8 seconds to 2.3 seconds. The model was announced on the Qwen large-model WeChat account.
Alibaba’s Qwen Releases Qwen3.8-LiveTranslate Simultaneous Interpretation Model
The notable element is the latency cut, since simultaneous interpretation is judged less on raw accuracy than on whether output keeps pace with speech. A drop of half a second per character moves the model closer to conversational viability, which matters for real-time settings such as meetings, media and cross-border commerce. Whether such gains hold outside controlled benchmarks, and how the Interleave architecture scales to other language pairs, is the open question.
Generated by AI for reference only.
Share on WeChat
Open WeChat → Scan → then tap "…" to send to a chat or Moments.
Tap "…" in the top-right corner to send to a chat or share to Moments.