llama.cpp 在 GPU offloading 的 MoE 测试里,E 核比 P 核更快

dir3ctly · reddit · 2026-07-25

原文链接 →

「Infra」频道最新

更多「Infra」频道 AI 资讯 →