ZeroHour
AI model

Qwen3.8-Omni-Flash

2 mentions in 7 days · 2 in 30 days · 2 total · first seen · last

Timeline

Alibaba Qwen Releases Qwen3.8-Omni-Flash: A 1M-Context Omni-Modal Model Built Around Agentic Audio-Video Understanding and Tool Use

Alibaba's Qwen team launched Qwen3.8-Omni-Flash, an API-only omni-modal model with 1M-token context and agentic audio-video understanding and tool use.

Qwen3.8-Omni-Flash accepts text, images, audio, and video and returns text, built on the Qwen3.8-Flash-Next architecture with a 1M-token context window and thinking enabled by default. Qwen reports a 25%+ average improvement over Qwen3.5-Omni-Plus across 29 evaluations, with OmniVideoBench rising from 63.4 to 67.8 while using about 45.7% fewer tokens via coarse-to-fine agentic perception. It is hosted on QwenCloud, Model Studio, and Qwen Studio at $0.15/$0.47 per 1M input/output tokens; no open weights were released, but Qwen open-sourced Qwen-MM-Plugins under Apache-2.0.

Alibaba releases Qwen 3.8 Omni Flash

Alibaba's Qwen team releases Qwen 3.8 Omni Flash, a fast omni-modal model announced on the official Qwen blog.

Alibaba's Qwen team announced the release of Qwen 3.8 Omni Flash via its official blog. The model is an omni-modal 'Flash' variant, positioned for fast, efficient multimodal inference. The announcement surfaced on Hacker News with 45 points and 7 comments; detailed parameter counts and benchmarks were not provided in the surfaced text.

Hacker News · AIupdated · 7h agofirst · 17h agoModel release 3 sourcesHN 45↑ · 7 comments

Appears with

Entities are extracted by the model from each article. Watching an entity keeps it in this browser only (no account); the watchlist page and dashboard alerts use it.