SyncAI.news, a Varaisys broadcasting
Qwen3.8-Omni-Flash undercuts Google's Gemini Flash pricing while matching its multimodal benchmarks
MB

Matthias Bastian

· 1 min read

BusinessThe Decoder

Qwen3.8-Omni-Flash undercuts Google's Gemini Flash pricing while matching its multimodal benchmarks

Qwen3.8-Omni-Flash is Qwen's first multimodal model built for AI agents. It processes audio and video together, draws conclusions, and uses tools on its own to edit vlogs, translate short videos, or summarize movies. The context window spans one million tokens. On audio-video tasks, Qwen says it comes close to matching Gemini 3.8 Flash.

API pricing sits at $0.15 per million input tokens and $0.47 per million output tokens. Qwen estimates audio input at under $0.01 per hour, while 720p video with audio at one frame per second runs about $0.20, not counting response costs. For comparison, Gemini 3.8 Flash charges $0.75 for input and $3.75 for output per million tokens at its introductory rate, with prices set to double on January 1, 2027.

The model is available through Qwen Studio, Qwen Cloud, and the API. The open-source Qwen-MM-Plugins add video editing, speaker recognition, PDF video notes, and reusable workflows to agents like Claude Code, Gemini CLI, and Qwen Code. Qwen-Live Harness enables real-time interaction using a camera and microphone.

AI News Without the Hype – Curated by Humans

Subscribe now

Original source

This story was published by The Decoder and written by Matthias Bastian. SyncAI.news shows a preview; the complete article is on the publisher's site.

Read the full story on the-decoder.com

Similar News