Local desktop AI girlfriend demo fits Whisper, llama.cpp and Qwen3-TTS into 15GB VRAM
xiaosun86 · x · 2026-07-26
A local-only AI girlfriend/desktop companion demo combines speech-to-speech components without using the cloud.
The cited setup uses Silero VAD v5 for voice activity detection, Whisper for STT, llama.cpp for the LLM, and Qwen3-TTS for speech synthesis. The creator says all four models fit into 15 GB of VRAM and can be hot-swapped freely on a fully local stack built on an open-source speech-to-speech project.
Related event: Fully Local AI Girlfriend Voice Demo Runs on 15GB VRAM(2 posts)→
More from Fun
- Interactive Video Generation in Any Art Style, Pure JS Coded by Claude — ctjlewis · 2026-09-23
- Claude wrote an entire song purely in code, no Suno involved — ctjlewis · 2026-09-23
- Economist worries AI detectors discriminate against aggressively literary French prose — paulnovosad · 2026-09-23
- When scientists can't find trends, they plot the logarithm of all variables — burny_tech · 2026-09-23
- ML author Burkov zings AI hype: 'visionary' claims vs Theranos founder in jail — burkov · 2026-09-23
- Your AI agent ran for 12 hours — but did anyone tell it the brief changed at hour 2? — HaktanSuren · 2026-09-23