Small business seeks local model advice beyond Qwen on ZGX Nano AI station
Maschinhunt · reddit · 2026-09-16
A Reddit user is building a local AI setup for a small business on a ZGX Nano G1N AI station, currently running Qwen 3.6 35B A3B (nvfp4) with 40-50GB RAM to spare. Main tasks are document processing, everyday Q&A and coding; they want something more precise and are considering Gemma 4 or Qwen 3.8 27B. MCP with SearXNG web search is already set up, and they also ask whether any model has RAG built in or whether RAG must be added separately in Open WebUI.
More from Infra
- Ben Bajarin doubles down on Credo: photonics-plus-SerDes vertical integration drives growth — BenBajarin · 2026-09-16
- Ocean Protocol launches hourly dedicated-GPU inference, H200 from $2.16/hour — w1kke · 2026-09-16
- NVIDIA Vera CPU completes agentic task lifecycles 1.64x faster than x86, Signal65 finds — ryanshrout · 2026-09-16
- Run your AI agents from anywhere with Tailscale and a simple PWA — johnlindquist · 2026-09-16
- DeepSeek V4.1 Flash Keeps Timing Out on 2-Hour Agentic Benchmarks, Author Shares Failure Logs — sebnadeau · 2026-09-16
- 800 VDC is the likely endgame for datacenter power as AI racks hit 120–300+ kW — BenBajarin · 2026-09-16