微软论文:蒸馏轨迹技能可省 2.7-6 倍输出 token,抵上推理模式

rohanpaul_ai · x · 2026-09-07

微软研究者在 arXiv 发布论文《Reason Wide, Not Deep: Amortizing the Reasoning Premium into Distilled Skills》,提出把昂贵的测试时推理成本“摊销”为可复用的技能规则。

所属事件:微软论文:蒸馏技能卡可省数倍token(2 条相关)→

原文链接 →

「编程与Agent」频道最新

更多「编程与Agent」频道 AI 资讯 →