Inception 发布扩散 LLM Mercury 2.5
Inception Labs 发布扩散式语言模型 Mercury 2.5,称智能较上代提升 40%,为目前训练过的最大扩散语言模型,在常见 NVIDIA GPU 上可达约 1107 tokens/秒,支持 26 万 token 上下文。模型已通过官方 API 及 OpenRouter、Baseten 提供,新账号赠 1 亿 tokens,发布期 8 折。落地方面,OpenCall 用其运行实时语音 Agent,p99 延迟降至 1 秒内、可压至 200ms。
2026-09-09 ~ 2026-09-09 · 4 条相关
- Inception 发布 Mercury 2.5:最强扩散 LLM,1107 tokens/秒、26 万上下文 — StefanoErmon · 2026-09-09
- OpenCall 用 Mercury 2.5 跑实时语音 Agent,p99 延迟降至 1 秒 — StefanoErmon · 2026-09-09
- 扩散模型 Mercury 2.5 上线:发布价 8 折,语音延迟压到 200ms 内 — StefanoErmon · 2026-09-09
另有 1 条近重复转述:timshi_ai