Alibaba ships Qwen3.8-Omni-Flash to watch, listen and call tools

Published Sep 17th, 2026, 9:02pm CT — corroborating coverage of Qwen team launch via Qianwen AI Platform, September 18th local time.

Alibaba is positioning Qwen3.8-Omni-Flash as a multimodal analysis and orchestration layer for agent developers, pairing a 1M-token context window with function calling, web search and discounted cached inputs.

According to Alibaba Cloud’s model documentation, Qwen3.8-Omni-Flash accepts text, images, audio and video through the Chat Completions and Responses APIs, with a 1 million-token context window. Its native output is text. Finished media and other work beyond analysis require external tools coordinated through the model’s function-calling capabilities.

Alibaba is pairing the model with Qwen-MM-Plugins for multimodal agent workflows and Qwen-Live Harness for continuous audiovisual interaction. Production documentation updated September 18th lists availability in Beijing, Singapore, Hong Kong, Tokyo, Frankfurt and Virginia.

Pricing and economics

Alibaba Cloud published international pricing at 0.016 per million cache-hit input tokens and $0.47 per million output tokens. Alibaba claims API cost per hour of audio input fell by more than 98% and audiovisual input by more than 93% versus prior models (company-reported calculation).

Benchmark claims

Alibaba claims an average improvement of more than 25% over Qwen3.5-Omni-Plus across 29 evaluations, with gains of 36.5 points on WildClawBench-MM and 22.3 points on AgenticVBench, alongside a 69.6 score on UniClawBench. Audio performance reportedly exceeded Gemini 3.8 Flash overall; audiovisual performance came close to Google’s model. Parameter count and training details not published.

Position in Qwen portfolio

Qwen3.8-Omni-Flash fills a lower-cost sensory layer for applications processing meetings, recordings, screens and video, complementing Qwen3.8-2.4T-A95B (open-weight MoE) and Qwen3.8-Max (coding/enterprise agents).

Primary announcement URL: https://qwen.ai/blog?id=qwen3.8-omni-flash