PEGAPOLL · NEWS

Qwen3.8-Omni-Flash matches Gemini's multimodal benchmarks at a fraction of the API price

The Decoder (AI) · 2026-09-19
🤖 AI-generated content — The title and summary were produced automatically by artificial intelligence, without human editorial review.

Qwen has released Qwen3.8-Omni-Flash, its first multimodal model built for AI agents, which processes audio and video together, draws conclusions, and uses tools on its own to edit vlogs, translate clips, or summarize movies. With a one-million-token context window, Qwen says it nearly matches Google's Gemini 3.8 Flash on audio-video benchmarks. Its API costs $0.15 per million input tokens and $0.47 per million output tokens, versus Gemini 3.8 Flash's introductory rate of $0.75 input and $3.75 output per million tokens, a price that is set to double on January 1, 2027. The model is available via Qwen Studio, Qwen Cloud, and the API.

Continue in the app — vote & join in ➔
Source: The Decoder (AI) · via ahirlevel.hu