Model Behavior Hinges on Post-Training
gerardsans · x · 2026-07-19
The discussion emphasizes that behavioral differences in models stem not just from pre-training data and architecture, but also from post-training and RLHF. The author points out that while many labs share common training corpora, the exact composition is opaque; furthermore, the data and rules used during post-training are often kept strictly behind closed doors. Therefore, evaluating model performance requires looking beyond pretraining to include post-training and the alignment process.
Related event: Claude's Constitutional AI: Post-Training Dictates Model Behavior(2 posts)→
More from Research
- GigaChat Audio targets long-form audio grounding with timestamps across 120-minute inputs — ai-sage · 2026-07-21
- Paper models Transformer components as stochastic geometry and tests five architectures — Zhihua Liang · 2026-07-21
- LTX 2.3 LoRA demo changes a video’s camera angle — CQDSN · 2026-07-21
- OpenForecaster uses daily news to improve language-model forecasting — Cohere_Labs · 2026-07-21
- Baseten study finds new facts in LLM weights are fragile unless trained from many restatements — alex_verem · 2026-07-21
- uv-scripts/ocr returns to the top of Hugging Face datasets with a JSON model picker — vanstriendaniel · 2026-07-21