Voice + LLM + just-in-time UI is becoming the universal interface layer, argues developer
rschu · x · 2026-09-04
Developer rschu laid out the case for voice + LLMs as the future of user interfaces:
- Core idea: speak naturally → STT captures it → an LLM understands intent and acts, far closer to human interaction than precise command sequences
- Evidence: the GPT-6 Astra demo's conversational multimodal interaction, Omarchy treating dictation as first-class, and his daily use of the open-source STT tool Handy to turn any text field into a voice interface
- His practice: dictating long prompts; imperfect transcription is fine since LLMs handle missing words, accents, and incomplete sentences well
- Visual UI won't disappear but becomes just-in-time UI, adapting to the task at hand; keyboards and touch remain for cases where they make sense
- Conclusion: voice + LLM + JIT UI is becoming the interface layer across OSes, apps, cars, robots, and spatial computing
More from AGI Musings
- Sam Altman predicted in 2015 that people would fall in love with AI chatbots by 2026 — TheIshanGoswami · 2026-09-04
- DHH slams useless GDPR cookie banners, warns of what happens when governments define AI — Dan_Jeffries1 · 2026-09-04
- Expert explainer on Anthropic's protein binder campaign: design cost cut from $10k to ~$100 — AllThingsApx · 2026-09-04
- Paradigm 3: low-quality RL environments may explain reward hacking; EBR-bench shows humans beat AIs — gleech · 2026-09-04
- Zero failure rate on alignment evals is a red flag, warn safety researchers — connoraxiotes · 2026-09-04
- Skeptical take: OpenAI can't train large models, pivots to RL and inference — teortaxesTex · 2026-09-04