Exploring new post-training methods for LLMs beyond chat templates
davidad · x · 2026-08-24
Suggests that there are likely many interesting ways to post-train LLMs if one moves away from standard chat templating constraints.
More from Research
- DataSpace Benchmark: Strong LLMs Fail as Reliable Data Agents — rohanpaul_ai · 2026-08-24
- Engineering gas vesicles into CAR-T cells lets ultrasound track them inside living organs — NikoMcCarty · 2026-08-24
- Alibaba PAI open-sources unified ControlNet for MiniMax-H3: one 7GB checkpoint, five control modes — linoy_tsaban · 2026-08-24
- SparsePR: Training-Free Sparse Attention for Video Generation — TexasAMUniversity · 2026-08-24
- Pure reinforcement learning enables zero-shot robot transfer — chris_j_paxton · 2026-08-24
- Mathematicians have low costs for using AI to solve problems, not relying on big labs — littmath · 2026-08-24