← 模型目录

Qwen3.8-Omni-Flash

Alibaba / Qwen · 2026-09-18 · 类别未确认

官方发布来源 · 按需求选型 · 查看覆盖缺口 · 归档与转录记录

Qwen3.8-Omni-Flash

Qwen 团队发布的新一代原生全模态大模型,支持 1M 上下文,具备从多模态内容理解到主动规划、工具调用与长程音视频创作的 Agentic 交付能力;在 29 项基准中平均较前代提升 25%,并大幅降低音视频 API 成本。

输入模态
文本 / 图像 / 音频 / 视频
上下文
1M
参数
官方资料未说明
价格(每百万 tokens)
无此口径报价;不等于免费

本变体的评测证据

uniclawbench 69.6 模型 qwen3-8-omni-flash · 版本 未说明 · 指标 score · 单位 score 来源等级 A · vendor_reported · 核对日期 2026-09-22 · 距快照 12 天

归一化读数:不计算;样本及方差不完整,无法计算置信区间。

原文位置与完整协议

heading: Evaluations · table: Omnimodal Agent Benchmarks · row: UniClawBench · quote_snippet: UniClawBench | 69.6

{
  "harness": "UniClawBench Agent Harness",
  "tools": "audio-visual player + browser + terminal",
  "shots": null,
  "reasoning_effort": null,
  "temperature": null,
  "top_p": null,
  "token_budget": null,
  "turn_limit": null,
  "time_limit": null,
  "run_count": null,
  "aggregation": null,
  "judge": null
}

UniClawBench 原生全模态智能体基准评测。new-benchmark

打开官方来源

omnivideobench 67.8 模型 qwen3-8-omni-flash · 版本 Agentic Mode · 指标 accuracy · 单位 score 来源等级 A · vendor_reported · 核对日期 2026-09-22 · 距快照 12 天

归一化读数:不计算;样本及方差不完整,无法计算置信区间。

原文位置与完整协议

heading: Evaluations · table: OmniVideoBench Agentic vs Static · row: OmniVideoBench (Agentic) · quote_snippet: OmniVideoBench (Agentic) | 67.8

{
  "harness": "Qwen-MM-Plugins Agentic Video Harness",
  "tools": "video segment seek + frame sampling",
  "shots": null,
  "reasoning_effort": null,
  "temperature": null,
  "top_p": null,
  "token_budget": null,
  "turn_limit": null,
  "time_limit": null,
  "run_count": null,
  "aggregation": null,
  "judge": null
}

OmniVideoBench 在智能体主动跳转与片段抽取模式下取得 67.8,较静态单轮模式提升显著并减少 45.7% Token

打开官方来源

omnivideobench 63.4 模型 qwen3-8-omni-flash · 版本 Static Mode · 指标 accuracy · 单位 score 来源等级 A · vendor_reported · 核对日期 2026-09-22 · 距快照 12 天

归一化读数:不计算;样本及方差不完整,无法计算置信区间。

原文位置与完整协议

heading: Evaluations · table: OmniVideoBench Agentic vs Static · row: OmniVideoBench (Static) · quote_snippet: OmniVideoBench (Static) | 63.4

{
  "harness": "Standard Direct Evaluation",
  "tools": null,
  "shots": null,
  "reasoning_effort": null,
  "temperature": null,
  "top_p": null,
  "token_budget": null,
  "turn_limit": null,
  "time_limit": null,
  "run_count": null,
  "aggregation": null,
  "judge": null
}

OmniVideoBench 静态直接视频理解模式自报分 63.4

打开官方来源

wildclawbench-mm +36.5 模型 qwen3-8-omni-flash · 版本 未说明 · 指标 gain · 单位 score 来源等级 A · vendor_reported · 核对日期 2026-09-22 · 距快照 12 天

归一化读数:不计算;样本及方差不完整,无法计算置信区间。

原文位置与完整协议

heading: Evaluations · quote_snippet: WildClawBench-MM (+36.5 points)

{
  "harness": "WildClawBench-MM Harness",
  "tools": null,
  "shots": null,
  "reasoning_effort": null,
  "temperature": null,
  "top_p": null,
  "token_budget": null,
  "turn_limit": null,
  "time_limit": null,
  "run_count": null,
  "aggregation": null,
  "judge": null
}

音视频复杂交互任务基准 WildClawBench-MM 相较 Qwen3.5-Omni-Plus 提升 36.5 分。new-benchmark

打开官方来源

agenticvbench +22.3 模型 qwen3-8-omni-flash · 版本 未说明 · 指标 gain · 单位 score 来源等级 A · vendor_reported · 核对日期 2026-09-22 · 距快照 12 天

归一化读数:不计算;样本及方差不完整,无法计算置信区间。

原文位置与完整协议

heading: Evaluations · quote_snippet: AgenticVBench (+22.3 points)

{
  "harness": "AgenticVBench Suite",
  "tools": null,
  "shots": null,
  "reasoning_effort": null,
  "temperature": null,
  "top_p": null,
  "token_budget": null,
  "turn_limit": null,
  "time_limit": null,
  "run_count": null,
  "aggregation": null,
  "judge": null
}

长程视频智能体基准 AgenticVBench 相较前代提升 22.3 分。new-benchmark

打开官方来源