π UPDATE β August 18, 2026
Elon Musk confirmed this morning that Grok 4.6 has now claimed the top spot on a key benchmark, adding further weight to the model's competitive positioning at launch. Musk shared the result directly on X alongside a chart screenshot, though the specific benchmark was not named in the post. This follows Grok 4.6's already-strong 1753 ELO score and aggressive $2/M input pricing, suggesting xAI is pulling ahead on both performance and value metrics simultaneously.
![]()
@elonmusk Β· Aug 18, 2026 β "Grok 4.6 takes top spot on this benchmark"
π UPDATE β August 15, 2026
Elon Musk has announced that Grok 4.6 has successfully completed "The Gauntlet," a demanding performance benchmark, sharing a video clip as evidence. While xAI has not yet published detailed scoring breakdowns for this specific test, the announcement signals continued capability expansion beyond the 1753 ELO figure reported at launch. The Gauntlet is widely regarded as a rigorous multi-domain stress test for frontier AI models, making this a notable milestone for Grok 4.6's real-world performance credentials. π Full benchmark details and scoring methodology from xAI have not yet been officially documented.
π UPDATE β August 14, 2026
Elon Musk has clarified that Grok 4.6 is optimized specifically for the Grok Build harness, warning that the experience will be "significantly worse" without it. He's advising developers and benchmarkers to evaluate Grok 4.6 exclusively through Build β a notable caveat that may affect how third-party ELO and performance comparisons are interpreted. This suggests the 1753 ELO score and headline benchmarks are likely achieved within the Build harness environment, not in raw API or standalone usage.
![]()
@elonmusk Β· Aug 14, 2026 β "Grok 4.6 will work best with the Grok Build harness. The experience will be significantly worse without it, so best to evaluate using Build."
π UPDATE β August 13, 2026
xAI has confirmed that Grok 4.6 now supports Computer-Aided Design (CAD) tasks, extending the model's reach into professional engineering and technical workflows. The official @grok account posted a demo video showcasing the capability, which marks a notable expansion beyond text and code into structured, precision-driven design use cases. This positions Grok 4.6 as a more direct competitor in technical and industrial AI applications alongside its already competitive 1753 ELO benchmark score and $2/M input pricing.
π UPDATE β August 13, 2026
xAI has expanded Grok 4.6's availability to two additional platforms: Pi and OpenCode, broadening access beyond the initial launch. To help users hit the ground running, xAI also reset usage limits β users can claim a reset token directly from Settings in the Grok desktop or mobile app. The limit reset appears timed specifically to coincide with the 4.6 rollout, giving developers and power users extra headroom to test the new model.
@grok Β· Aug 13, 2026
"We've reset limits to help you keep building during the Grok 4.6 launch. Use a reset token from settings in Grok on desktop or mobile."
Grok 4.6 now available in Pi β via @SpaceXAI
π UPDATE β August 13, 2026
Real-world testing is reinforcing Grok 4.6's capabilities beyond the benchmark numbers. Tech creator @DirtyTesLa shared a demonstration where Grok 4.6 β working autonomously from a single prompt for 22 minutes β produced a project complete with custom shaders, a minimap, and a time-change feature. The extended autonomous work session highlights the model's ability to sustain complex, multi-step reasoning and code generation without additional user input. This kind of long-horizon task completion is emerging as one of Grok 4.6's standout real-world differentiators alongside its headline 1753 ELO score.
@DirtyTesLa Β· Aug 13, 2026
"Grok 4.6 is legit. Custom shaders, minimap, time change. And this is from one prompt, Grok worked for 22 minutes."
![]()
π UPDATE β August 12, 2026
Grok 4.6 has now claimed the #1 spot on the Databricks leaderboard, adding another top benchmark ranking to its already impressive 1753 ELO score on LMSYS Chatbot Arena. Elon Musk announced the milestone on X, sharing screenshots of the Databricks results. This makes Grok 4.6 the first xAI model to reach #1 on Databricks, further validating its position as a leading frontier model β and at half the price of key competitors.
Source: @elonmusk on X
xAI has officially released Grok 4.6, the next iteration of its flagship large language model, positioning it as a substantial jump over Grok 4.5 at the same API price. Elon Musk called the release a 'banger' on X, citing a 1753 ELO score and pricing that undercuts other frontier models by roughly half. The model is live today across Grok Build, Cursor, Grok Bot, and the xAI API, with a first-week promotion doubling included usage in Cursor and Grok Build.

What xAI Is Claiming
The official announcement thread from the xAI account frames Grok 4.6 as a step-change on three axes: intelligence, speed, and cost. According to xAI, the model 'delivers frontier intelligence and is a significant improvement over Grok 4.5 at the same price,' and 'can handle much more challenging tasks' than its predecessor. Musk followed up with a benchmark claim β a 1753 ELO score β and the marketing pitch of 'amazing bang for buck.'
According to independent reporting from OrcaRouter and kie.ai, Grok 4.6 is a 1.5-trillion-parameter model built on the same V9 foundation as Grok 4.5, but with a longer supplemental training run that layered in improved supervised fine-tuning and reinforcement learning on curated data. In other words: same base architecture, materially more training work on top.
Key Figures
| Metric | Grok 4.6 | Notes |
|---|---|---|
| API input pricing | $2 / M tokens | xAI states this is half the price of other frontier models |
| API output pricing | $6 / M tokens | Faster variant available at 2x this rate, per OrcaRouter |
| ELO score (xAI-cited) | 1753 | Cited by Musk; not yet independently verified on public arenas |
| Parameter count | 1.5 trillion | Same V9 base as Grok 4.5, per OrcaRouter reporting |
| Launch date | Aug 7, 2026 (rollout); public announcement Aug 12 | Hits target Musk set publicly on July 28 |
| Promo | 2x included usage | First week only, Cursor + Grok Build |

The Pricing Story Is the Real Headline
The ELO number will get the retweets, but the pricing is arguably the bigger competitive move. At $2 per million input tokens and $6 per million output tokens, Grok 4.6 sits well below the sticker price of comparable frontier tier offerings from OpenAI and Anthropic. xAI's messaging is explicit: 'half the price of other frontier models.' That framing is aimed squarely at developers and agent builders whose bills scale with token volume β the exact cohort xAI needs to win to catch up on API market share.
Pairing that pricing with a first-week 2x usage promotion inside Cursor β one of the most popular AI-native code editors β is a distribution play. Cursor's developer audience is where competitive AI model preferences get decided in real workflows, not on Twitter benchmarks. Doubling usage during launch week means Cursor users will actively test Grok 4.6 against whatever model they're currently defaulting to, at essentially no marginal cost.
The 1753 ELO Claim, With Caveats
Musk's 1753 ELO figure is worth flagging carefully. According to independent tracking on prediction market platforms like PredictionHunt, as of August 11 β one day before this announcement β public prediction markets were still re-evaluating the likelihood of Grok 4.6 hitting high ELO scores on platforms like LMSYS Chatbot Arena, and independent benchmarks had not yet published. The 1753 number is xAI-cited and appears to reflect internal or early third-party arena scoring. Public leaderboards typically take days to weeks to stabilize as vote volume accumulates.
Separately, xAI states Grok 4.6 matches GPT-5.6 Sol on the Artificial Analysis Intelligence Index with a composite score of 61, according to reporting from OfficeChai and kie.ai. That's a more structured benchmark and, if it holds up in independent testing, would put Grok 4.6 in the same tier as the current frontier β at half the price.

Where You Can Use It Today
Grok 4.6 is live as of the announcement across four primary surfaces: the xAI API, Grok Build (xAI's developer platform), Cursor (the AI coding editor), and Grok Bot on X. Third-party access is also live via OpenRouter, Vercel, and Cloudflare, according to xAI's launch materials β meaning developers who route requests through model marketplaces can switch to Grok 4.6 without changing providers.
Existing Grok subscribers on X should see the model rotate in automatically. Cursor users need to select Grok 4.6 from the model dropdown to trigger the 2x usage promo. API developers can point requests at the new model identifier immediately.
What This Means for the Tesla-Adjacent AI Story
For Tesla owners tracking xAI's arc, Grok 4.6 matters because xAI is now the AI supplier being wired into Tesla vehicles for in-car voice interaction and, longer term, into Optimus. A cheaper, faster, higher-scoring Grok makes the economics of shipping AI features across millions of Tesla endpoints materially better. Every cent of token cost gets multiplied by fleet scale.
The cadence xAI is signaling is aggressive. According to OrcaRouter, Grok 4.7 β a larger 2.1-trillion-parameter model β is expected within weeks of the 4.6 release, and Grok 5 is targeted before the end of 2026. If those dates hold, xAI is pushing a release tempo that few frontier labs have sustained.
What to Watch Next
- Independent ELO confirmation. LMSYS Chatbot Arena and Artificial Analysis will post community-verified scores within days. Whether the 1753 figure holds up publicly is the first credibility test.
- Cursor adoption data. If Grok 4.6 becomes a default in a meaningful share of Cursor sessions after the 2x promo ends, that's the strongest signal the pricing pitch is working.
- Grok 4.7 timing. A larger 2.1T-parameter model was reported as imminent. A shipping date confirms the aggressive release cadence.
- Tesla in-car integration. Watch for a Tesla OTA that references an updated Grok voice assistant model β the natural downstream beneficiary of a cheaper, faster frontier tier.
The next 72 hours of independent benchmark data will decide whether Grok 4.6 is genuinely competitive at the top of the frontier stack, or whether the ELO claim outpaces the reality. Either way, the pricing move is real, it's live today, and it puts direct cost pressure on every other frontier lab's API business.
Related Gear
Gear up your Tesla with tested, custom-fit BASENOR accessories β shop Tesla accessories β
Sources & reporting notes
The links below identify the material source records used for this report.
- @SpaceXAI on X (2026-08-12T15:32:11.000Z) β Direct source
- @SpaceXAI on X (2026-08-12T15:32:11.000Z) β Direct source
- @SpaceXAI on X (2026-08-12T15:32:12.000Z) β Direct source
- @michaelnicollsx on X (2026-08-12T15:36:34.000Z) β Direct source
- @elonmusk on X (2026-08-12T15:41:00.000Z) β Direct source
- @elonmusk on X (2026-08-12T15:42:25.000Z) β Direct source
- @elonmusk on X (2026-08-12T15:46:31.000Z) β Direct source
Source links are preserved as published or accessed. See our editorial standards and corrections policy.
The BASENOR Editorial Desk covers Tesla, SpaceX, and related technology, curating reporting from primary sources β official accounts, regulatory filings, and software release data. Every article passes source-record and fact-checking review before publication. About the newsroom.
This report was curated by the BASENOR Editorial Desk from the sources listed above. Read our editorial standards or email editorial@basenor.com to report an error.









