← 模型目录

Hunyuan T1 (混元-T1 正式版)

Tencent / 腾讯混元 · 2025-03-21 · 类别未确认

官方发布来源 · 按需求选型 · 查看覆盖缺口

Hunyuan T1 (混元-T1 正式版)

腾讯混元-T1 正式版定位为业内首个超大规模混合 Mamba 推理模型(TurboS 混合 Transformer-Mamba MoE 底座,96.7% 后训练算力用于强化学习)。已收录 14 项评测覆盖知识、数学、代码与中文能力:亮点 MMLU-Pro 87.2、AIME24 78.2。

输入模态
文本
上下文
官方资料未说明
参数
官方资料未说明
价格(每百万 tokens)
无此口径报价;不等于免费

本变体的评测证据

mmlu-pro 87.2 模型 hunyuan-t1 · 版本 未说明 · 指标 accuracy · 单位 percent 来源等级 A · vendor_reported · 核对日期 2026-09-01 · 距快照 33 天

归一化读数:不计算;样本及方差不完整,无法计算置信区间。

原文位置与完整协议

heading: 混元-T1: 强化学习驱动,业内首个超大规模混合Mamba推理模型正式发布! (README prose) · row: mmlu-pro · quote_snippet: MMLU-PRO 上 T1 仅次于 O1,高达 87.2 分

{
  "harness": null,
  "tools": null,
  "shots": null,
  "reasoning_effort": null,
  "temperature": null,
  "top_p": null,
  "token_budget": null,
  "turn_limit": null,
  "time_limit": null,
  "run_count": null,
  "aggregation": null,
  "judge": null
}

打开官方来源

gpqa 69.3 模型 hunyuan-t1 · 版本 Diamond · 指标 accuracy · 单位 percent 来源等级 A · vendor_reported · 核对日期 2026-09-01 · 距快照 33 天

归一化读数:不计算;样本及方差不完整,无法计算置信区间。

原文位置与完整协议

heading: 混元-T1: 强化学习驱动,业内首个超大规模混合Mamba推理模型正式发布! (README prose) · row: gpqa · quote_snippet: GPQA- diamond … T1 达到了 69.3 分

{
  "harness": null,
  "tools": null,
  "shots": null,
  "reasoning_effort": null,
  "temperature": null,
  "top_p": null,
  "token_budget": null,
  "turn_limit": null,
  "time_limit": null,
  "run_count": null,
  "aggregation": null,
  "judge": null
}

打开官方来源

lcb 64.9 模型 hunyuan-t1 · 版本 未说明 · 指标 accuracy · 单位 percent 来源等级 A · vendor_reported · 核对日期 2026-09-01 · 距快照 33 天

归一化读数:不计算;样本及方差不完整,无法计算置信区间。

原文位置与完整协议

heading: 混元-T1: 强化学习驱动,业内首个超大规模混合Mamba推理模型正式发布! (README prose) · row: lcb · quote_snippet: LiveCodeBench 的代码评测中,T1 达到了 64.9 分

{
  "harness": null,
  "tools": null,
  "shots": null,
  "reasoning_effort": null,
  "temperature": null,
  "top_p": null,
  "token_budget": null,
  "turn_limit": null,
  "time_limit": null,
  "run_count": null,
  "aggregation": null,
  "judge": null
}

打开官方来源

math500 96.2 模型 hunyuan-t1 · 版本 未说明 · 指标 accuracy · 单位 percent 来源等级 A · vendor_reported · 核对日期 2026-09-01 · 距快照 33 天

归一化读数:不计算;样本及方差不完整,无法计算置信区间。

原文位置与完整协议

heading: 混元-T1: 强化学习驱动,业内首个超大规模混合Mamba推理模型正式发布! (README prose) · row: math500 · quote_snippet: MATH-500 上,取得了 96.2 分

{
  "harness": null,
  "tools": null,
  "shots": null,
  "reasoning_effort": null,
  "temperature": null,
  "top_p": null,
  "token_budget": null,
  "turn_limit": null,
  "time_limit": null,
  "run_count": null,
  "aggregation": null,
  "judge": null
}

打开官方来源

arenahard 91.9 模型 hunyuan-t1 · 版本 未说明 · 指标 win_or_tie_rate · 单位 percent 来源等级 A · vendor_reported · 核对日期 2026-09-01 · 距快照 33 天

归一化读数:不计算;样本及方差不完整,无法计算置信区间。

原文位置与完整协议

heading: 混元-T1: 强化学习驱动,业内首个超大规模混合Mamba推理模型正式发布! (README prose) · row: arenahard · quote_snippet: ArenaHard 任务中,T1 拿下了 91.9 分

{
  "harness": null,
  "tools": null,
  "shots": null,
  "reasoning_effort": null,
  "temperature": null,
  "top_p": null,
  "token_budget": null,
  "turn_limit": null,
  "time_limit": null,
  "run_count": null,
  "aggregation": null,
  "judge": null
}

打开官方来源

chinese-simpleqa 67.9 模型 hunyuan-t1 · 版本 未说明 · 指标 accuracy · 单位 percent 来源等级 A · vendor_reported · 核对日期 2026-09-01 · 距快照 33 天

归一化读数:不计算;样本及方差不完整,无法计算置信区间。

原文位置与完整协议

heading: README benchmark table (image) · row: chinese-simpleqa · figure: data/model-releases/official/_archive/tencent/hunyuan-t1/images/01-bench-table.png

{
  "harness": null,
  "tools": null,
  "shots": null,
  "reasoning_effort": null,
  "temperature": null,
  "top_p": null,
  "token_budget": null,
  "turn_limit": null,
  "time_limit": null,
  "run_count": null,
  "aggregation": null,
  "judge": null
}

Value read from the page appendix image and confirmed by re-reading the archived image with the Read tool on 2026-09-01 (data/model-releases/official/_archive/tencent/hunyuan-t1/images/01-bench-table.png). Page prints the row under the Hunyuan T1 column; protocol fields transcribed from the notes block printed beneath the same table.

打开官方来源

drop 93.1 模型 hunyuan-t1 · 版本 F1 · 指标 f1 · 单位 percent 来源等级 A · vendor_reported · 核对日期 2026-09-01 · 距快照 33 天

归一化读数:不计算;样本及方差不完整,无法计算置信区间。

原文位置与完整协议

heading: README benchmark table (image) · row: drop · figure: data/model-releases/official/_archive/tencent/hunyuan-t1/images/01-bench-table.png

{
  "harness": null,
  "tools": null,
  "shots": null,
  "reasoning_effort": null,
  "temperature": null,
  "top_p": null,
  "token_budget": null,
  "turn_limit": null,
  "time_limit": null,
  "run_count": null,
  "aggregation": null,
  "judge": null
}

Value read from the page appendix image and confirmed by re-reading the archived image with the Read tool on 2026-09-01 (data/model-releases/official/_archive/tencent/hunyuan-t1/images/01-bench-table.png). Page prints the row under the Hunyuan T1 column; protocol fields transcribed from the notes block printed beneath the same table.

打开官方来源

zebralogic 79.6 模型 hunyuan-t1 · 版本 未说明 · 指标 accuracy · 单位 percent 来源等级 A · vendor_reported · 核对日期 2026-09-01 · 距快照 33 天

归一化读数:不计算;样本及方差不完整,无法计算置信区间。

原文位置与完整协议

heading: README benchmark table (image) · row: zebralogic · figure: data/model-releases/official/_archive/tencent/hunyuan-t1/images/01-bench-table.png

{
  "harness": null,
  "tools": null,
  "shots": null,
  "reasoning_effort": null,
  "temperature": null,
  "top_p": null,
  "token_budget": null,
  "turn_limit": null,
  "time_limit": null,
  "run_count": null,
  "aggregation": null,
  "judge": null
}

Value read from the page appendix image and confirmed by re-reading the archived image with the Read tool on 2026-09-01 (data/model-releases/official/_archive/tencent/hunyuan-t1/images/01-bench-table.png). Page prints the row under the Hunyuan T1 column; protocol fields transcribed from the notes block printed beneath the same table.

打开官方来源

aime24 78.2 模型 hunyuan-t1 · 版本 未说明 · 指标 accuracy · 单位 percent 来源等级 A · vendor_reported · 核对日期 2026-09-01 · 距快照 33 天

归一化读数:不计算;样本及方差不完整,无法计算置信区间。

原文位置与完整协议

heading: README benchmark table (image) · row: aime24 · figure: data/model-releases/official/_archive/tencent/hunyuan-t1/images/01-bench-table.png

{
  "harness": null,
  "tools": null,
  "shots": null,
  "reasoning_effort": null,
  "temperature": null,
  "top_p": null,
  "token_budget": null,
  "turn_limit": null,
  "time_limit": null,
  "run_count": null,
  "aggregation": null,
  "judge": null
}

Value read from the page appendix image and confirmed by re-reading the archived image with the Read tool on 2026-09-01 (data/model-releases/official/_archive/tencent/hunyuan-t1/images/01-bench-table.png). Page prints the row under the Hunyuan T1 column; protocol fields transcribed from the notes block printed beneath the same table.

打开官方来源

ceval 91.8 模型 hunyuan-t1 · 版本 未说明 · 指标 accuracy · 单位 percent 来源等级 A · vendor_reported · 核对日期 2026-09-01 · 距快照 33 天

归一化读数:不计算;样本及方差不完整,无法计算置信区间。

原文位置与完整协议

heading: README benchmark table (image) · row: ceval · figure: data/model-releases/official/_archive/tencent/hunyuan-t1/images/01-bench-table.png

{
  "harness": null,
  "tools": null,
  "shots": null,
  "reasoning_effort": null,
  "temperature": null,
  "top_p": null,
  "token_budget": null,
  "turn_limit": null,
  "time_limit": null,
  "run_count": null,
  "aggregation": null,
  "judge": null
}

Value read from the page appendix image and confirmed by re-reading the archived image with the Read tool on 2026-09-01 (data/model-releases/official/_archive/tencent/hunyuan-t1/images/01-bench-table.png). Page prints the row under the Hunyuan T1 column; protocol fields transcribed from the notes block printed beneath the same table.

打开官方来源

cmmlu 90 模型 hunyuan-t1 · 版本 未说明 · 指标 accuracy · 单位 percent 来源等级 A · vendor_reported · 核对日期 2026-09-01 · 距快照 33 天

归一化读数:不计算;样本及方差不完整,无法计算置信区间。

原文位置与完整协议

heading: README benchmark table (image) · row: cmmlu · figure: data/model-releases/official/_archive/tencent/hunyuan-t1/images/01-bench-table.png

{
  "harness": null,
  "tools": null,
  "shots": null,
  "reasoning_effort": null,
  "temperature": null,
  "top_p": null,
  "token_budget": null,
  "turn_limit": null,
  "time_limit": null,
  "run_count": null,
  "aggregation": null,
  "judge": null
}

Value read from the page appendix image and confirmed by re-reading the archived image with the Read tool on 2026-09-01 (data/model-releases/official/_archive/tencent/hunyuan-t1/images/01-bench-table.png). Page prints the row under the Hunyuan T1 column; protocol fields transcribed from the notes block printed beneath the same table.

打开官方来源

cfbench 81 模型 hunyuan-t1 · 版本 未说明 · 指标 accuracy · 单位 percent 来源等级 A · vendor_reported · 核对日期 2026-09-01 · 距快照 33 天

归一化读数:不计算;样本及方差不完整,无法计算置信区间。

原文位置与完整协议

heading: README benchmark table (image) · row: cfbench · figure: data/model-releases/official/_archive/tencent/hunyuan-t1/images/01-bench-table.png

{
  "harness": null,
  "tools": null,
  "shots": null,
  "reasoning_effort": null,
  "temperature": null,
  "top_p": null,
  "token_budget": null,
  "turn_limit": null,
  "time_limit": null,
  "run_count": null,
  "aggregation": null,
  "judge": null
}

new-benchmark: cfbench not yet in data/benchmarks/. Value read from the page appendix image and confirmed by re-reading the archived image with the Read tool on 2026-09-01 (data/model-releases/official/_archive/tencent/hunyuan-t1/images/01-bench-table.png). Page prints the row under the Hunyuan T1 column; protocol fields transcribed from the notes block printed beneath the same table.

打开官方来源

cello 76.4 模型 hunyuan-t1 · 版本 未说明 · 指标 accuracy · 单位 percent 来源等级 A · vendor_reported · 核对日期 2026-09-01 · 距快照 33 天

归一化读数:不计算;样本及方差不完整,无法计算置信区间。

原文位置与完整协议

heading: README benchmark table (image) · row: cello · figure: data/model-releases/official/_archive/tencent/hunyuan-t1/images/01-bench-table.png

{
  "harness": null,
  "tools": null,
  "shots": null,
  "reasoning_effort": null,
  "temperature": null,
  "top_p": null,
  "token_budget": null,
  "turn_limit": null,
  "time_limit": null,
  "run_count": null,
  "aggregation": null,
  "judge": null
}

new-benchmark: cello not yet in data/benchmarks/. Value read from the page appendix image and confirmed by re-reading the archived image with the Read tool on 2026-09-01 (data/model-releases/official/_archive/tencent/hunyuan-t1/images/01-bench-table.png). Page prints the row under the Hunyuan T1 column; protocol fields transcribed from the notes block printed beneath the same table.

打开官方来源

t-eval 68.8 模型 hunyuan-t1 · 版本 未说明 · 指标 accuracy · 单位 percent 来源等级 A · vendor_reported · 核对日期 2026-09-01 · 距快照 33 天

归一化读数:不计算;样本及方差不完整,无法计算置信区间。

原文位置与完整协议

heading: README benchmark table (image) · row: t-eval · figure: data/model-releases/official/_archive/tencent/hunyuan-t1/images/01-bench-table.png

{
  "harness": null,
  "tools": null,
  "shots": null,
  "reasoning_effort": null,
  "temperature": null,
  "top_p": null,
  "token_budget": null,
  "turn_limit": null,
  "time_limit": null,
  "run_count": null,
  "aggregation": null,
  "judge": null
}

new-benchmark: t-eval not yet in data/benchmarks/. Value read from the page appendix image and confirmed by re-reading the archived image with the Read tool on 2026-09-01 (data/model-releases/official/_archive/tencent/hunyuan-t1/images/01-bench-table.png). Page prints the row under the Hunyuan T1 column; protocol fields transcribed from the notes block printed beneath the same table.

打开官方来源