← 模型目录

Gemini 3.8 Live / Gemini 3.8 Live Extended Thinking

Google DeepMind / Gemini · 2026-09-15 · 类别未确认

官方发布来源 · 按需求选型 · 查看覆盖缺口 · 归档与转录记录

Gemini 3.8 Live

Google 最新原生语音到语音实时对话模型,专为低延迟自然人机交互优化,支持异步工具调用、实时视觉对齐与 97+ 语言无缝流转切换。

输入模态
文本 / 音频 / 图像 / 视频
上下文
1M
参数
官方资料未说明
价格(每百万 tokens)
USD 输入 未说明 / 输出 未说明 · 音频输入 $0.005/min ($0.84/hr),音频输出 $0.018/min

本变体的评测证据

尚无对应评测记录。缺少证据不代表能力为零。

Gemini 3.8 Live Extended Thinking

面向高复杂度任务的双工对话推理模型,支持边对话边在后台执行多步深度思考与工具调度,登顶 Artificial Analysis 语音对语音质量指数首位 (82.6)。

输入模态
文本 / 音频 / 图像 / 视频
上下文
1M
参数
官方资料未说明
价格(每百万 tokens)
USD 输入 未说明 / 输出 未说明 · 音频输入 $0.005/min,音频输出 $0.018/min

本变体的评测证据

aa-speech-to-speech 82.6 模型 gemini-3-8-live-extended-thinking · 版本 未说明 · 指标 quality_index · 单位 index 来源等级 A · third_party_reported · 核对日期 2026-09-22 · 距快照 12 天

归一化读数:不计算;样本及方差不完整,无法计算置信区间。

原文位置与完整协议

heading: Evaluations · table: Speech-to-Speech Quality · row: Artificial Analysis Speech to Speech Quality Index · quote_snippet: Artificial Analysis Speech-to-Speech Quality Index | 82.6 (#1 Overall)

{
  "harness": "Artificial Analysis Speech to Speech Suite",
  "tools": null,
  "shots": null,
  "reasoning_effort": "high",
  "temperature": null,
  "top_p": null,
  "token_budget": null,
  "turn_limit": null,
  "time_limit": null,
  "run_count": null,
  "aggregation": null,
  "judge": null
}

Artificial Analysis Speech to Speech Quality Index 榜首 82.6 分。new-benchmark

打开官方来源

tau-voice 68.6% 模型 gemini-3-8-live-extended-thinking · 版本 未说明 · 指标 task_completion · 单位 percent 来源等级 A · vendor_reported · 核对日期 2026-09-22 · 距快照 12 天

归一化读数:不计算;样本及方差不完整,无法计算置信区间。

原文位置与完整协议

heading: Evaluations · table: Voice Agent Task Completion · row: τ-Voice (Tau-Voice) · quote_snippet: τ-Voice (Tau-Voice) Agentic Benchmark | 68.6%

{
  "harness": "Tau-Voice Agent Harness",
  "tools": "voice telephony API + async tool execution",
  "shots": null,
  "reasoning_effort": "high",
  "temperature": null,
  "top_p": null,
  "token_budget": null,
  "turn_limit": null,
  "time_limit": null,
  "run_count": null,
  "aggregation": null,
  "judge": null
}

Tau-Voice 语音自主任务完成率 68.6%。new-benchmark

打开官方来源

tau-voice-banking 35.1% 模型 gemini-3-8-live-extended-thinking · 版本 未说明 · 指标 task_completion · 单位 percent 来源等级 A · vendor_reported · 核对日期 2026-09-22 · 距快照 12 天

归一化读数:不计算;样本及方差不完整,无法计算置信区间。

原文位置与完整协议

heading: Evaluations · table: Voice Agent Task Completion · row: Sierra τ-Voice-Banking · quote_snippet: Sierra τ-Voice-Banking Benchmark | 35.1%

{
  "harness": "Sierra Voice-Banking Harness",
  "tools": "banking transactional API",
  "shots": null,
  "reasoning_effort": "high",
  "temperature": null,
  "top_p": null,
  "token_budget": null,
  "turn_limit": null,
  "time_limit": null,
  "run_count": null,
  "aggregation": null,
  "judge": null
}

Sierra 银行电话客服高合规自主任务完成率 35.1%。new-benchmark

打开官方来源

bigbench-audio 97.7% 模型 gemini-3-8-live-extended-thinking · 版本 未说明 · 指标 accuracy · 单位 percent 来源等级 A · vendor_reported · 核对日期 2026-09-22 · 距快照 12 天

归一化读数:不计算;样本及方差不完整,无法计算置信区间。

原文位置与完整协议

heading: Evaluations · table: Audio Understanding & Reasoning · row: Big Bench Audio · quote_snippet: Big Bench Audio | 97.7%

{
  "harness": "BigBench Audio Evaluation Suite",
  "tools": null,
  "shots": null,
  "reasoning_effort": "high",
  "temperature": null,
  "top_p": null,
  "token_budget": null,
  "turn_limit": null,
  "time_limit": null,
  "run_count": null,
  "aggregation": null,
  "judge": null
}

Big Bench Audio 复杂音频推理与理解准确率 97.7%。new-benchmark

打开官方来源