Expressiveness in voice AI should come from conversation context, not a style setting
deepgramscott · x · 2026-08-18
Deepgram's Scott argues that the meaning of a sentence rarely lives in the words themselves: "Sure." can mean agreement, annoyance, disbelief, or "stop talking" depending on context. So expressiveness in voice AI shouldn't be a style setting but a product of conversational context—if a user has been frustrated for five minutes, the next response shouldn't suddenly sound upbeat.
More from Multimodal
- Open Source Single-File Video Tool Optimized for MiniMax H3 Ref2V — bstr3k · 2026-08-19
- Helios: Generative video shifts focus from quality to real-time efficiency and interaction — AI Engineer · 2026-08-19
- MiniMax H3 video generation demo with last frame enhanced by Krea 2 — takayatodoroki · 2026-08-19
- MiniMax H3 speed optimization: 10s video in 10 mins on 5070 Ti — Tight_Commercial7 · 2026-08-19
- LFM2.5-VL-3B Demo: Auto-Maps Skin Regions and Measures — pcuenq · 2026-08-19
- Seedance 2.5 Creates Complete AI Influencer Vlog in 30 Seconds — aftahi_ai · 2026-08-19