Radium racked its own GPUs to undercut OpenAI and Anthropic pricing — small devs still won't switch

No_Raspberry7273 · reddit · 2026-09-10

Radium, a small North American team, bought and racked GPUs to serve tuned open-weight models behind OpenAI- and Anthropic-compatible endpoints, arguing that classification, extraction and formatting tasks shouldn't pay frontier-model prices. The pitch lands with enterprises — several run production workloads with them — but solo devs create API keys and never return. The author rules out trust, switching cost (just a baseurl change) and docs, and asks the community: what's the actual bar for moving workloads to a smaller inference provider, and where has "OpenAI-compatible" broken on you?

Original post →

More from Venture

Venture channel →