Best local uncensored vision+text models for an RTX 5070 Ti 16GB setup?

Mystvearn_ · reddit · 2026-09-12

A Reddit user wants a fully local, uncensored vision+text workflow on an RTX 5070 Ti (16GB VRAM) with 64GB DDR4, mimicking cloud models like Gemini and Claude for image description and prompt brainstorming. They ask which open-weight multimodal models, quantization levels, and software stacks (Ollama, LM Studio) best fit the 16GB constraint.

Original post →

More from Infra

Infra channel →