One prompt burned 88% of weekly quota: Astra user hits a sub-agent usage drain
NanoIsAMeme · reddit · 2026-09-10
An OpenAI Astra user reports his weekly quota plunged from 93% remaining to 5% after a single prompt. He uses Astra high for planning specs and Astra low for implementation, and had explicitly banned sub-agents in his agents.md to avoid usage drain. But when he asked the model to review an open-source project called Omniagent in a concurrent session, usage instantly dropped 80%. The takeaway: asking models to analyze agent-type projects can silently trigger heavy sub-agent fan-out, even against explicit configuration.
More from Models
- InclusionAI's free Ling 3.0 Flash Sante: a 124B MoE medical model with 262K context on OpenRouter — FellMentKE · 2026-09-10
- DeepSeek V4.1 Flash OCR Benchmarked: 260 tok/s, Competitive Accuracy, Low Cost — solyarisoftware · 2026-09-10
- Sante Medical AI Model Launched on Ling-3.0-flash: Focus on Reasoning, Safety, Evidence — FellMentKE · 2026-09-10
- Sante Medical AI Model Launched on Ling-3.0-flash: Focus on Reasoning, Safety, Evidence — FellMentKE · 2026-09-10
- Sante Medical AI Model Launched on Ling-3.0-flash: Focus on Reasoning, Safety, Evidence — FellMentKE · 2026-09-10
- Yoav Goldberg: 880K GPU-Hours vs a Few Hundred Prompted LLM Hours for Same Math Result — yoavgo · 2026-09-10