Gemini 5.6 Luna Cached Input Price Drops to $0.02/M Tokens
rudrank · x · 2026-07-31
A developer reported that the cached input price for the Gemini 5.6 Luna model has dropped to $0.02/M tokens. This highly competitive pricing will significantly reduce inference costs for developers handling long contexts or high-frequency API calls.
Related event: Llama 3.1 and Gemini Cut Cached Input Prices to $0.02(2 posts)→
More from Infra
- The Cost of 'Good Enough' Data: Why Modern Architectures Fail at Scale — craigmullins · 2026-07-31
- Big Tech AI spending tops $1 trillion, FT reports — gaganghotra_ · 2026-07-31
- OpenAI Slashes GPT-5.6 Prices by 80%, Inference Cost Drops 2000x Annually — Latent Space · 2026-07-31
- Revisiting Lossy Verification in Speculative Decoding: Mechanisms and Failure Modes — Tianyu Wang · 2026-07-31
- Energy Consumption: Single AI Prompt vs Agentic Workflow Differs by 100,000x — AndyMasley · 2026-07-31
- The AI Trade Runs on Borrowed Money, and Lenders Are Repricing It — haipothetical · 2026-07-31