llama.cpp 报错:Ling 3.0 Flash 模型加载出现大量 Unused Tensor

Debianreiser69 · reddit · 2026-08-22

用户在使用 llama.cpp 加载 Ling-3.0-flash-Q6K 模型(分片 GGUF)时遇到报错,日志显示 blk.42 中的多个张量(如 attnq.weight, ffngateexps.weight)被标记为“未使用”并被忽略。用户表示该错误也出现在腾讯 Hunyuan 等其他模型上。目前仅展示了日志,尚无解决方案。

原文链接 →

「Infra」频道最新

更多「Infra」频道 AI 资讯 →