Using Base Models for More Natural User Simulations
_ddjohnson · x · 2026-09-01
To make simulated users more realistic, the system uses a pretraining-only base model to generate user messages, with an assistant-trained "pilot" selecting the best candidate. This approach increases naturalness without compromising controllability.
Related event: OpenAI Uses Simulated Users for Multi-Turn LLM Evaluation(2 posts)→
More from Research
- NeurReps 2026 CFP: Symmetry and Geometry in Neural Representations — fatihdin4en · 2026-09-01
- Dan Luu on why software slowness is a choice, analyzing latency costs and optimization — JeremyCMorgan · 2026-09-01
- Scholar calls out LLM gibberish: reviewing papers and replies is now a waste of time — thegautamkamath · 2026-09-01
- Paper analyzes reasoning models like o1 and DeepSeek R1, probing CoT data contamination — rao2z · 2026-09-01
- Qdrant's Sept 17 stream: token-native storage claims 10-100x faster reads — qdrant_engine · 2026-09-01
- Qdrant Event Preview: Benchmarks on Hybrid Search Tuning Parameters — qdrant_engine · 2026-09-01