TWT 方法将 ViT 冗余层折叠为单层,算力近乎减半精度几乎不降

KyeGomezB · x · 2026-09-26

挪威研究团队(Oslo 大学医院/奥斯陆大学等)提出 Transformer-Within-Transformer(TWT),一种事后(post-hoc)模型压缩方法。

原文链接 →

「模型」频道最新

更多「模型」频道 AI 资讯 →