Missing API for general real-time LLM agents: AsyncLLM preprint sparks interface debate
phill1992 · reddit · 2026-09-30
A Reddit discussion asks whether real-time LLM interaction can be generalized: models already handle mid-thought stimuli in specific cases (robotics agents, GPT-4o voice, video calls, mid-turn steering in Codex), but there's no general API for building real-time agents the way developers write custom MCP tools.
Points raised: vendor real-time APIs (Gemini Live, OpenAI Realtime) target voice/coding niches; compelling use cases include coding agents you steer mid-debug, deep-research agents that react to feedback mid-search, and voice-driven assistants. The closest general approach is the AsyncLLM preprint, where programmers write asyncio coroutines with shared-memory blocks for agent communication — but it still requires building a low-level inference pipeline. The open question: what would the right interface look like for a general async-agent API that OpenAI/Anthropic could expose for feeding live event streams?
More from coding & agent
- ModRetro and OpenAI bring Codex-built games to the Chromatic handheld — Dimillian · 2026-09-30
- First DeepSeek Harness plugin ships: Qiaomu AI RSS brings AI news and podcasts into your coding agent — vista8 · 2026-09-30
- DoorDash launches corporate ordering MCP and CLI so AI agents can place orders, Vercel and Cognition early adopters — Scobleizer · 2026-09-30
- YC launches Video Arena, a free arena where AI models compete to code your video prompts — ycombinator · 2026-09-30
- openharness: One Tmux-Based TUI to Command Every Coding Agent on Every Machine — dee_hw · 2026-09-30
- Your MCP anonymizer is self-defeating if it takes code as an argument — Aggressive-Course-24 · 2026-09-30