Dev discovers vLLM accepts input embeds, speculates finetuning tokens are underused

cephaloform · x · 2026-10-09

开发者 @cephaloform 表示此前因 vLLM 接口变动太快而回避使用,现在才发现 vLLM 支持传入 input embeds,打算实际尝试。他在讨论中提出一个猜想:社区可能严重低估了 prompt tuning 类微调 token(以及 DreamBooth 式高显存配置)的潜力,如果推理框架普遍支持接受 inputsids/embeds,结合进化算法或许能打开新玩法。

Original post →

More from Infra

Infra channel →