Mixing RTX 3060 and 5060 Ti for local 30B model inference?
MkGod · reddit · 2026-08-18
A user discusses the feasibility of combining an RTX 5060 Ti (16GB) with a used RTX 3060 (12GB) to achieve 28GB VRAM for running 30B+ models locally. Key discussion points include:
- Performance Bottlenecks: The impact of the slower Ampere card and PCIe 3.0 x4 lane on overall throughput and memory bandwidth.
- VRAM Utility: Practical gains of 16GB over 12GB in terms of KV cache headroom and quantization levels.
- Cost Analysis: Comparing this setup against purchasing an RTX 4060 Ti or a second 5060 Ti.
More from Embodied
- Robot breaks speed record, smashes electrical box — skdh · 2026-08-18
- Humanoid Robots Now Handle Packages With Human-Like Coordination — CurieuxExplorer · 2026-08-18
- Robotics for Agriculture and Resilience: Ruggedize Event Draws 250+ Deep Techies — Scobleizer · 2026-08-18
- One night of vibe coding: open-source Reachy robot runs 100% local with Qwen brain — Scobleizer · 2026-08-18
- Giving the Bristlebot a Brain: First Agentic Micro-Robot That Senses, Predicts, and Acts — maier_ak · 2026-08-18
- Stanford releases VISTA semantic navigation repo — rsasaki0109 · 2026-08-18