FULL STORY

Grok 4.7: From Leak to Launch

Grok 4.7 briefly surfaced in OpenCode Zen's catalog ahead of launch, and xAI officially released it days later, touting double the speed at half the price with major benchmark gains.

2026-09-21 ~ 2026-09-22 · 3 episodes · 36 posts

Episode 1 · Grok 4.7 briefly surfaces on OpenCode Zen, launch rumored imminent (2026-09-21, 5 posts)

On September 21, xAI's next-generation model Grok 4.7 appeared to be nearing release: several X users spotted grok-4.7 briefly listed in the OpenCode Zen model directory before it was removed, widely read as a sign that a launch was imminent. That same day, another user claimed Grok 4.7 would drop that afternoon, sharing specific specs and pricing details. xAI has not officially confirmed anything so far.

Confirmed

  • grok-4.7 did appear in the OpenCode Zen model directory and was subsequently removed. @kimmonismus, @CurieuxExplorer, and @socialwithaayan all reported the sighting; @CurieuxExplorer joked that if it's on the menu, it's already in the kitchen.
  • xAI has made no official confirmation about the Grok 4.7 release as of now.

Unconfirmed

  • @ns123abc posted that Grok 4.7 would launch early that afternoon, with screenshots as evidence, but there is no official backing.
  • @realsohamparekh claimed Grok 4.7 would launch today with a 500K tokens context window, multimodal support, and pricing similar to Grok 4.6. The claim comes from a personal account and has not been verified by xAI.
  • Both @kimmonismus and @socialwithaayan predicted this would be a "big model release week": after xAI, OpenAI (Dev Day) and Anthropic could make moves in quick succession — though this is personal speculation.

Why it matters

  • A model name briefly appearing in the OpenCode Zen directory is typically seen as a precursor to an official rollout, so the community views this incident as a strong signal that xAI's new model is about to launch.
  • If the rumors of a 500K context window, multimodal capabilities, and pricing on par with Grok 4.6 hold true, Grok 4.7 would compete directly with rivals on long-context performance and value for money — but everything hinges on xAI's official announcement.

Episode 2 · Grok 4.7 rolls out at same price with big gains on long-horizon tasks (2026-09-21, 5 posts)

xAI has released Grok 4.7, with multiple community users confirming the new model is live and reportedly appearing in the xAI API, corroborating the launch news. Taken together, the leaks suggest the new version delivers significant benchmark and capability improvements at unchanged pricing—the most notable story of this release cycle.

Confirmed

  • Grok 4.7 has shipped and is available: users including @DanielLockyer confirmed the new version is live, and @iamfakhrealam reports, citing LuminaBench, that Grok 4.7 has appeared in the xAI API.
  • Pricing stays flat versus 4.6: $2/M input, $6/M output, plus a double-speed tier at twice the price (via @markk).
  • Upgrade direction: a larger base model with longer RL training; officials say the new model can persist longer on hard multi-hour tasks and check its results more carefully (via @nimaowji, @markk).
  • Benchmark performance: Terminal-Bench improved from 20.3% to 38.0% (via @markk, noted as unverified in the post).

Not Yet Confirmed

  • Most posts are leaks from third-party accounts, and early posts (@markk, @koltregaskes) explicitly noted they were unverified by official xAI channels; later community confirmations and the API sighting boost credibility, but a full official statement from xAI and detailed benchmark data are still pending.

Why It Matters

  • Faster performance at the same price, plus a near-doubling on Terminal-Bench, directly improves Grok's cost-performance competitiveness in agent and terminal task scenarios.
  • The emphasis on multi-hour long-horizon tasks and self-checking signals that xAI is positioning agent scenarios as its main competitive front.

Episode 3 · xAI Launches Grok 4.7, Its Strongest Coding Model Yet (2026-09-22, 26 posts)

xAI (SpaceXAI) officially released Grok 4.7 on September 22, positioning it as its most powerful model for coding and knowledge work. The company claims it runs twice as fast as comparable models at half the price—actual pricing stays level with Grok 4.6 at $2/million input tokens and $6/million output tokens. The notable jump in benchmark scores is the highlight of this release.

Confirmed

  • Grok 4.7 is built on a new, larger base model with extended reinforcement learning for multi-step tasks lasting hours; it is better at self-verification and long-context management, and natively understands the Grok Bot system
  • CursorBench score reaches 46.3%, surpassing GPT-5.6 Sol's 41.7%; Terminal-Bench leaps from the previous generation's 20.3% to 38%; Harvey legal benchmarks also improved substantially
  • Pricing is unchanged from Grok 4.6: $2/million input tokens, $6/million output tokens, which the company says is a fraction of what comparable competitors charge

Not Yet Confirmed

  • Third-party commentary from blogger XFreeze claims Grok 4.7 shows a massive leap on multi-hour office tasks (AA Briefcase), surpassing GPT-6 Astra and approaching Fable 5.1—this is personal interpretation; refer to official benchmark tables for exact scores

Why It Matters

  • Major benchmark gains at the same price and speed, combined with lower pricing than competitors, position Grok 4.7 to shake up the price-performance landscape for coding and knowledge-work models; the strengthened ability to handle long multi-step tasks with self-verification also targets real needs of agentic office workflows

6 more related posts →