---
title: "Alibaba releases Qwen3.8-Omni-Flash, a 1M-context omni-modal model"
url: https://www.parallelquant.com/posts/alibaba-releases-qwen3-8-omni-flash-a-1m-context-omni-modal-model-bba097
source_name: "MarkTechPost"
source_url: https://www.marktechpost.com/2026/09/18/alibaba-qwen-releases-qwen3-8-omni-flash/
published: 2026-09-18T08:40:37.000Z
topics: ["llms", "open source"]
publisher: "Parallel Quant"
---

# Alibaba releases Qwen3.8-Omni-Flash, a 1M-context omni-modal model

*2026-09-18 · Source: [MarkTechPost](https://www.marktechpost.com/2026/09/18/alibaba-qwen-releases-qwen3-8-omni-flash/)*

Alibaba's Qwen team released Qwen3.8-Omni-Flash, an omni-modal model with a 1 million token context window that processes audio and video, plans tasks, and calls tools. The company reports it uses about 45.7% fewer tokens on the OmniVideoBench benchmark than prior approaches.

**Why it matters:** Large context windows paired with native audio-video understanding and tool use push open-weight models closer to matching proprietary multimodal agents, and the reported token-efficiency gain matters directly for cost at scale for anyone running long video or audio agentic workloads.

**Topics:** llms, open source

---
Read the original: https://www.marktechpost.com/2026/09/18/alibaba-qwen-releases-qwen3-8-omni-flash/
Canonical: https://www.parallelquant.com/posts/alibaba-releases-qwen3-8-omni-flash-a-1m-context-omni-modal-model-bba097
Published by Parallel Quant — https://www.parallelquant.com
