Free Qwen3.8-27B endpoint launched: 262K context, vision, and tool calls

victormustar · x · 2026-08-15

Developer victormustar deployed a free public endpoint for Qwen3.8-27B that requires no token and is fully OpenAI-compatible.

Running on a single H200 GPU, the endpoint supports vision input, tool calling, and adjustable reasoning intensity. It features a 262K-token context window. The deployment costs roughly $5/hour, supports 50 concurrent requests, and achieves 60 tokens/s. The service will remain online for at least 72 hours.

Original post →

More from Infra

Infra channel →