Gemini 3.8 Live / Gemini 3.8 Live Extended Thinking
Google DeepMind / Gemini · 2026-09-15 · 类别未确认
官方发布来源 · 按需求选型 · 查看覆盖缺口 · 归档与转录记录
Gemini 3.8 Live
Google 最新原生语音到语音实时对话模型,专为低延迟自然人机交互优化,支持异步工具调用、实时视觉对齐与 97+ 语言无缝流转切换。
- 输入模态
- 文本 / 音频 / 图像 / 视频
- 上下文
- 1M
- 参数
- 官方资料未说明
- 价格(每百万 tokens)
- USD 输入 未说明 / 输出 未说明 · 音频输入 $0.005/min ($0.84/hr),音频输出 $0.018/min
本变体的评测证据
尚无对应评测记录。缺少证据不代表能力为零。
Gemini 3.8 Live Extended Thinking
面向高复杂度任务的双工对话推理模型,支持边对话边在后台执行多步深度思考与工具调度,登顶 Artificial Analysis 语音对语音质量指数首位 (82.6)。
- 输入模态
- 文本 / 音频 / 图像 / 视频
- 上下文
- 1M
- 参数
- 官方资料未说明
- 价格(每百万 tokens)
- USD 输入 未说明 / 输出 未说明 · 音频输入 $0.005/min,音频输出 $0.018/min
本变体的评测证据
归一化读数:不计算;样本及方差不完整,无法计算置信区间。
原文位置与完整协议
heading: Evaluations · table: Speech-to-Speech Quality · row: Artificial Analysis Speech to Speech Quality Index · quote_snippet: Artificial Analysis Speech-to-Speech Quality Index | 82.6 (#1 Overall)
{
"harness": "Artificial Analysis Speech to Speech Suite",
"tools": null,
"shots": null,
"reasoning_effort": "high",
"temperature": null,
"top_p": null,
"token_budget": null,
"turn_limit": null,
"time_limit": null,
"run_count": null,
"aggregation": null,
"judge": null
}Artificial Analysis Speech to Speech Quality Index 榜首 82.6 分。new-benchmark
归一化读数:不计算;样本及方差不完整,无法计算置信区间。
原文位置与完整协议
heading: Evaluations · table: Voice Agent Task Completion · row: τ-Voice (Tau-Voice) · quote_snippet: τ-Voice (Tau-Voice) Agentic Benchmark | 68.6%
{
"harness": "Tau-Voice Agent Harness",
"tools": "voice telephony API + async tool execution",
"shots": null,
"reasoning_effort": "high",
"temperature": null,
"top_p": null,
"token_budget": null,
"turn_limit": null,
"time_limit": null,
"run_count": null,
"aggregation": null,
"judge": null
}Tau-Voice 语音自主任务完成率 68.6%。new-benchmark
归一化读数:不计算;样本及方差不完整,无法计算置信区间。
原文位置与完整协议
heading: Evaluations · table: Voice Agent Task Completion · row: Sierra τ-Voice-Banking · quote_snippet: Sierra τ-Voice-Banking Benchmark | 35.1%
{
"harness": "Sierra Voice-Banking Harness",
"tools": "banking transactional API",
"shots": null,
"reasoning_effort": "high",
"temperature": null,
"top_p": null,
"token_budget": null,
"turn_limit": null,
"time_limit": null,
"run_count": null,
"aggregation": null,
"judge": null
}Sierra 银行电话客服高合规自主任务完成率 35.1%。new-benchmark
归一化读数:不计算;样本及方差不完整,无法计算置信区间。
原文位置与完整协议
heading: Evaluations · table: Audio Understanding & Reasoning · row: Big Bench Audio · quote_snippet: Big Bench Audio | 97.7%
{
"harness": "BigBench Audio Evaluation Suite",
"tools": null,
"shots": null,
"reasoning_effort": "high",
"temperature": null,
"top_p": null,
"token_budget": null,
"turn_limit": null,
"time_limit": null,
"run_count": null,
"aggregation": null,
"judge": null
}Big Bench Audio 复杂音频推理与理解准确率 97.7%。new-benchmark