Inception 发布扩散 LLM Mercury 2.5:1107 tokens/秒,对标 GPT-5.6 Luna
thione · x · 2026-09-15
- Inception(CEO 为 Stefano Ermon)发布 Mercury 2.5,称其为目前最强、规模最大的扩散式语言模型,智能水平较 Mercury 2 提升 40%,可对标 GPT-5.6 Luna (Low)、Gemini 3.5 Flash-Lite、Claude Haiku 4.5 等低成本前沿模型。
- 关键指标:1107 tokens/秒(常规 NVIDIA GPU)、26 万 token 上下文、定价 $0.20/百万输入、$0.75/百万输出;发布期 8 折($0.04/$0.15)。
- 支持可调推理、并行工具调用、schema 对齐 JSON;官方称自 Mercury 2 以来使用量增长超一个数量级,新模型基于生产失败案例改进 evals 与训练。
- 同帖提及 Google DeepMind 推出 AlphaGenome Atlas(详见另一条)。
「模型」频道最新
- Agent Arena 榜单:DeepSeek V4.1 Flash 以 $0.06/任务挤进 Pareto 前沿 — arena · 2026-09-15
- 23 天没开 Claude Code:博主称 Codex 配开源模型更顺手 — Yuchenj_UW · 2026-09-15
- 深度估计模型 Marigold-V2 冲上 Hugging Face 热榜 — toshas · 2026-09-15
- 曝 OpenAI 雇数百合同工人工审读 ChatGPT 对话 — The Decoder · 2026-09-15
- Hugging Face 团队梳理:哪些开源 LLM 最适合端侧推理 — NielsRogge · 2026-09-15
- 用户实测指认:推理商 nahcrof 疑用 tokenizer 不符的模型冒充 Kimi K3 — xeophon · 2026-09-15