Post-Cutoff.com
  1. Home
  2. Timeline
  3. 2026
  4. Alibaba releases Qwen3.8-Omni-Flash, an omnimodal agent…

Alibaba releases Qwen3.8-Omni-Flash, an omnimodal agent model with 1M context and 93-98% cheaper audio/video input

★★★after cutoffmodel-releaseAlibabaQwenconfidence: high

Alibaba's Qwen team launched Qwen3.8-Omni-Flash, a native omnimodal model (text, image, audio and video in; 1M-token context) built for audio/video agent work such as video editing, film commentary and meeting summaries. Qwen reports a >25% average gain over Qwen3.5-Omni-Plus on 29 evaluations, audio performance above Gemini 3.8 Flash, and API prices per hour of audio (audio-visual) input more than 98% (93%) lower. It also open-sourced the Qwen-Live Harness and expanded Qwen-MM-Plugins.

Key facts

What happened

Qwen3.8-Omni-Flash is the Qwen3.8-generation successor to the Qwen Omni line. Qwen positions it as a step from understanding audio and video to acting on them: planning tasks, calling tools and delivering finished work (edited videos, music videos, commentary, PDF notes from tutorials) on its own. It also improves long-audio understanding and multi-speaker recognition. Alongside the model, Qwen open-sourced Qwen-Live Harness, a runtime for continuous real-time omnimodal interaction, and added plugins such as Video2Note. It is closed-weights and served on Alibaba Cloud Model Studio and the Qianwen app.

Why it matters

The model sells audio and video understanding at Flash-tier prices and competes directly with Gemini 3.8 Flash on the multimodal agents that Chinese and US labs are both racing to ship.

All benchmarks and price comparisons are Qwen's own. The exact release day is uncertain (Sept 14 to 18).

Changelog

  • 2026-09-30: created (the model file existed but there was no timeline entry)

Models

Related events

  1. Alibaba launches Qwen-Audio-3.1 five-model voice stack and cuts audio API prices up to 95% ★★★
  2. Apsara 2026: Alibaba says Qwen 4 is in training, targets 5-10T-parameter Qwen 4.5/5, reports self-improvement runs and unveils Zhenwu V900 chip ★★★
  3. Qwen3.8-Flash-Next: 125B MoE with only 6B active previews Qwen 4 architecture ★★★

Sources (4)

id: 2026-09-18-qwen3-8-omni-flash · updated 2026-09-30 · open in the interactive timeline