Muse Realtime Avatar
Meta / Llama · 2026-09-23 · 类别未确认
官方发布来源 · 按需求选型 · 查看覆盖缺口 · 归档与转录记录
Muse Realtime Avatar
Meta 音频驱动的实时数字人模型,基于因果 Diffusion Transformer 架构,通过自强迫分布匹配蒸馏将推理评估从 120 次减少到 2 次(60x 加速),支持 448x768 25fps 视频流式生成,端到端延迟约 870ms,真人双向盲测相对 Runway Characters 偏好度 78%:22%、相对 HeyGen LiveAvatar 偏好度 88%:12%。
- 输入模态
- 音频 / 图像 / 视频
- 上下文
- 官方资料未说明
- 参数
- 官方资料未说明
- 价格(每百万 tokens)
- 无此口径报价;不等于免费
本变体的评测证据
归一化读数:不计算;样本及方差不完整,无法计算置信区间。
原文位置与完整协议
heading: Live evaluation · row: Live Evaluation vs Runway Characters · quote_snippet: Live evaluation chart showing Muse Realtime Avatar preferred over Runway Characters by 78% to 22%
{
"harness": null,
"tools": null,
"shots": null,
"reasoning_effort": null,
"temperature": null,
"top_p": null,
"token_budget": null,
"turn_limit": null,
"time_limit": null,
"run_count": null,
"aggregation": null,
"judge": null
}new-benchmark: 2-3 分钟实时对话盲测,整体偏好度 78% vs 22%(Runway Characters);涵盖面部表现力、动作自然度、唇形同步、角色保持等维度。
归一化读数:不计算;样本及方差不完整,无法计算置信区间。
原文位置与完整协议
heading: Live evaluation · row: Live Evaluation vs HeyGen LiveAvatar · quote_snippet: and over HeyGen LiveAvatar by 88% to 12%
{
"harness": null,
"tools": null,
"shots": null,
"reasoning_effort": null,
"temperature": null,
"top_p": null,
"token_budget": null,
"turn_limit": null,
"time_limit": null,
"run_count": null,
"aggregation": null,
"judge": null
}new-benchmark: 2-3 分钟实时对话盲测,整体偏好度 88% vs 12%(HeyGen LiveAvatar)。
归一化读数:不计算;样本及方差不完整,无法计算置信区间。
原文位置与完整协议
heading: Live evaluation · row: End-of-turn to first-byte latency · quote_snippet: Muse Realtime Avatar streams 448x768 portrait video at 25 frames per second with approximately 870 ms of latency as measured from the end of a user’s turn to when they receive the first byte of the synchronized AV stream
{
"harness": null,
"tools": null,
"shots": null,
"reasoning_effort": null,
"temperature": null,
"top_p": null,
"token_budget": null,
"turn_limit": null,
"time_limit": null,
"run_count": null,
"aggregation": null,
"judge": null
}new-benchmark: 用户发言结束到接收到首字节同步音视频流的端到端延迟约为 870ms。