全球大模型进展新闻浏览站
中文头版English
输入关键词,快速查找已抓取新闻。

今日版 / 2026年9月27日星期日

limbo logolimbo

数据更新时间

9月19日 14:30

启用来源

17

抓取状态

真实抓取

研究进展The Decoder

Qwen3.8-Omni-Flash undercuts Google's Gemini Flash pricing while matching its multimod...

摘要

这条新闻已抓取入库,中文摘要正在生成中。你可以先打开原文查看完整内容。

原文译文

以下为抓取到的原文内容译文,已统一为站内阅读格式。

阅读原文

Qwen3.8-Omni-Flash undercuts Google's Gemini Flash pricing while matching its multimodal benchmarks

Qwen3.8-Omni-Flash undercuts Google's Gemini Flash pricing while matching its multimodal benchmarks

Qwen3.8-Omni-Flash is Qwen's first multimodal model built for AI agents. It processes audio and video together, draws conclusions, and uses tools on its own to edit vlogs, translate short videos, or summarize movies. The context window spans one million tokens. On audio-video tasks, Qwen says it comes close to matching Gemini 3.8 Flash.

Qwen 3.8 Omni Flash performs on par with Gemini Flash 3.8 in multimodal benchmarks but is much more affordable. | Image: Qwen

API pricing sits at $0.15 per million input tokens and $0.47 per million output tokens. Qwen estimates audio input at under $0.01 per hour, while 720p video with audio at one frame per second runs about $0.20, not counting response costs. For comparison, Gemini 3.8 Flash charges $0.75 for input and $3.75 for output per million tokens at its introductory rate, with prices set to double on January 1, 2027.

The model is available throughQwen Studio,Qwen Cloud, and theAPI. The open-sourceQwen-MM-Pluginsadd video editing, speaker recognition, PDF video notes, and reusable workflows to agents like Claude Code, Gemini CLI, and Qwen Code.Qwen-Live Harnessenables real-time interaction using a camera and microphone. Ad Ad

AI News Without the Hype – Curated by Humans

Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section.

来源地区

Europe

热度分

81

分类

研究进展

语言

en