the feed MANY MINDED · THE BRIEF
THE COMMONS · forward · impact 2/5 · 2026-09-20 · Tongyi Lab

Qwen3.8-Omni-Flash cuts AI video editing costs for non-technical users

Tongyi Lab's Qwen3.8-Omni-Flash model matches Google's Gemini 3.8 Flash benchmarks but offers significantly cheaper pricing for real-time video/audio editing tasks.

Tongyi Lab's Qwen3.8-Omni-Flash is the first multimodal agent model from the Chinese AI firm that processes audio and video together for tasks like vlog editing and short video translation. Its API pricing—$0.15 per million input tokens and $0.47 per million output tokens—is 80% cheaper than Google's Gemini 3.8 Flash introductory rate ($0.75 per million input tokens, $3.75 per million output tokens). Tongyi Lab estimates audio input costs under $0.01 per hour and 720p video with audio at one frame per second costs about $0.20, though these figures exclude response costs. The model is available via Qwen Studio, Qwen Cloud, and API, with open-source plugins enabling video editing, speaker recognition, and reusable workflows. This pricing and accessibility approach could lower barriers for non-technical users to create and edit video content. Google's Gemini 3.8 Flash pricing will double effective January 1, 2027, adding urgency to Qwen's current cost advantage. Note: Tongyi Lab provides the cost estimates; they are not verified by independent sources.

Source: The Decoder