Local voice assistant on Qwen3.5 4B with tool calling, no dedicated GPU
Leading_Yogurt7025 · reddit · 2026-10-11
A Reddit user shares a fully local voice assistant setup that runs without a dedicated GPU:
- LLM: Qwen3.5 4B with skills/tool calling
- STT: Voxtral; TTS: Pocket
- Entire pipeline runs locally; a longer YouTube demo video is linked
More from Infra
- MCIO cables extend PCIe at half-slot bandwidth, retimers fix signal loss — TheZachMueller · 2026-10-11
- llama.cpp merges probabilistic MTP decoding, +14% speedup on prose generation — Dreeew84 · 2026-10-11
- Motherboard to switchboard won't populate your PCIe slots properly — TheZachMueller · 2026-10-11
- Cloudflare: sites blocked 13.47% of AI bot requests in Q3, more than double a year ago — YvesMulkers · 2026-10-11
- Final vllm-radiance Build Enables Multi-Agent on One R9700 GPU — KriptacMessage · 2026-10-11
- Leak: OpenAI compute-constrained while Anthropic scales limits via SpaceX Colossus — mark_k · 2026-10-11