Karpathy’s voice-first prompting trick turns messy rambling into cleaner prompts
dair_ai · x · 2026-07-22
The post highlights a multimodal prompting pattern from Karpathy: speak out a long, messy stream of thought instead of typing it, and let the model reconstruct the intent more cleanly.
The quoted reply adds that voice can be mixed with other modalities for richer prompting, and says a recorded demo was used to show how well this works in practice.
Related event: Karpathy Recommends 10-Minute Voice Rambling for LLM Context(17 posts)→
More from Multimodal
- 3D ResNet Paper Crosses 3,000 Citations Eight Years After CVPR 2018 — HirokatuKataoka · 2026-09-11
- Non-coder builds full-featured Android ComfyUI client with ChatGPT, submits to Google Play — ComfierUI · 2026-09-11
- RunningHub open-sources H3Lightning, speeding up MiniMax H3 video generation 12x — 智东西 · 2026-09-11
- FastH3-Live hits 22fps: acceleration node benchmarks and the --vram-headroom trick — spartong945 · 2026-09-11
- Midjourney style code share: --sref 2912175708 — tisch_eins · 2026-09-11
- Astra storyboards plus Minimax H3 per-shot generation boost video success rates — Hailuo_AI · 2026-09-11