September 18, 2026 · MarkTechPost
Alibaba releases Qwen3.8-Omni-Flash, a 1M-context omni-modal model
Alibaba's Qwen team released Qwen3.8-Omni-Flash, an omni-modal model with a 1 million token context window that processes audio and video, plans tasks, and calls tools. The company reports it uses about 45.7% fewer tokens on the OmniVideoBench benchmark than prior approaches.
Why it matters: Large context windows paired with native audio-video understanding and tool use push open-weight models closer to matching proprietary multimodal agents, and the reported token-efficiency gain matters directly for cost at scale for anyone running long video or audio agentic workloads.