Anime, manga, and games, with a take · A Yukimedia publication

← all stories otherrelease 1 sources · 1h ago ·

Alibaba Adds Qwen3.8-Omni-Flash To Qwen Model Line

Alibaba is pricing audio and video input down hard while closing the benchmark gap to Gemini 3.8 Flash, which puts the cost of agent workflows built on omni-modal models under pressure.

Reporting from 1 source: GIGAZINE.

Alibaba Adds Qwen3.8-Omni-Flash To Qwen Model Line

Alibaba released Qwen3.8-Omni-Flash, a native omni-modal model in its Qwen line. It supports a 1-million-token context window and keeps text performance equal to text-only models of the same size. Across 29 benchmarks, its average score rose more than 25% over Qwen3.5-Omni-Plus, while API pricing for audio input fell more than 98% and audio-video input more than 93%.

The model runs on Qianwen AI and ships with speech recognition for 74 languages and speech generation for 29. Its reasoning scores include 82.7 on LongAudioSpan, 63.4 on OmniVideoBench, 28.2 on OmniCap-IF, and 89.7 on AliMeeting. On agent, coding, and long-horizon tasks it scored 71.0 on WildClawBench-MM, up 36.5 over Qwen3.5-Omni-Plus, and 69.6 on UniClawBench. Alibaba says it surpassed Gemini 3.8 Flash on multiple tests.

Synthesized by Yomimono from the 1 cited source below, including Japanese-language reporting where cited, then editorially reviewed before publishing.

Sources