Research cut-off: August 24, 2026. Grok 4.6 has been available since August 12, 2026. Current first-party SpaceXAI documentation confirms the model ID grok-4.6, a 500,000-token context window, text and image input, text output, and base pricing of $2 per million input tokens and $6 per million output tokens. Cached input is $0.50 per million, with higher input rates above 200,000 prompt tokens. [S43][S44]
MarketWatch’s reported five-point improvement over Grok 4.5 on the Artificial Analysis Intelligence Index remains a qualified external report, not a benchmark we reproduced ourselves. [S01] The final parameter count and a directly reproducible independent 4.6 score are still not cleanly documented. That separation is more useful for a Mac decision than a precise-looking but unsupported spec sheet.
The six things to know first
- Grok 4.6 launched on August 12, 2026 and is listed in current SpaceXAI documentation. [S10][S43]
- MarketWatch reports a five-point improvement over Grok 4.5 on the Artificial Analysis Intelligence Index. [S01]
- SpaceXAI lists $2/M input, $0.50/M cached input and $6/M output, with higher input rates above 200K prompt context. [S43][S44]
- The official model ID is
grok-4.6, with Responses API and OpenAI-compatible examples. [S43][S44] - The official context window is 500,000 tokens, with text/image input and text output. [S43][S44]
- The final parameter count and a directly reproducible independent 4.6 score remain open.
Why Grok 4.6 matters for agents
SpaceXAI made Grok 4.5 available on the xAI API on July 8, with the detailed product launch following in mid-July. The model targets coding, agentic work and knowledge tasks; its official launch material emphasized multi-step software-engineering reinforcement learning, tool-heavy workloads and high serving speed. [S07] Less than a month later, 4.6 extends that cadence.
The strongest post-launch signal is not a vendor adjective. It is the reported movement in independent evaluation. MarketWatch says Grok 4.6 gained five points versus 4.5 on Artificial Analysis’ Intelligence Index and moved back toward the frontier group. [S01]
That matters most for workloads in which a model must keep making useful decisions after dozens of tool calls rather than merely produce a good one-shot answer.
The benchmark result needs one important qualification
The accessible Artificial Analysis pages give Grok 4.5 a score of 54 in the July 8 snapshot. [S26][S27][S29] A five-point increase would arithmetically imply about 59.
That does not make “Grok 4.6 scores 59” a safe headline.
There are at least three reasons to keep the wording conservative:
- the press report may refer to a newer leaderboard snapshot;
- aggregate benchmark methodology can be revised;
- the exact 4.6 model/settings entry was not directly recoverable from the accessible AA snapshot used for this research.
The defensible sentence is therefore: MarketWatch reports that Grok 4.6 improved by five points over Grok 4.5 on the Artificial Analysis Intelligence Index. [S01]
Pricing: the base tier is now documented
SpaceXAI’s current Grok 4.6 model page lists $2/M input tokens, $0.50/M cached input tokens and $6/M output tokens. Prompts above 200,000 tokens move to a higher input tier, so the regional pricing table should be checked before production use. [S43][S44]
Using only those reported rates:
| Workload | Input | Output | Formula | Estimated list cost |
|---|---|---|---|---|
| small | 0.5M | 0.1M | 0.5×$2 + 0.1×$6 | $1.60 |
| medium | 10M | 2M | 10×$2 + 2×$6 | $32 |
| heavy | 100M | 20M | 100×$2 + 20×$6 | $320 |
These figures use the official base tier. They deliberately exclude cache hits, long-context surcharges, tool charges, priority-processing charges and reseller markups. SpaceXAI documents higher input rates above 200,000 prompt tokens; use the live regional table for a production estimate. [S43][S44]
For an agent product, that distinction is material. A cheap token price can be offset by more retries, longer outputs or weak cache reuse.
1.5 trillion or 2 trillion parameters?
If you searched for Grok 4.6 before launch, you probably saw 2T.
IT之家, ComputerBase and other July coverage described a two-trillion-parameter successor to Grok 4.5. [S32][S33] Kie.ai built a detailed pre-launch explainer around the same number. [S40]
Later July reporting changed the story. EdgeX, The Economic Times and RB.ru attributed a 1.5T figure to Elon Musk and described improved supervised fine-tuning and reinforcement-learning work, with a larger Grok 4.7 expected later. [S30][S31][S34]
The sound conclusion is not to pick whichever number appears in more articles. Both clusters trace back to pre-launch statements, and the accessible post-launch SpaceXAI model documentation does not settle the final production architecture.
So the spec table should say:
Parameter count: not yet verified in an accessible post-launch primary model card.
That single line is more useful than a confident but stale number.
Context, modalities and tools: now documented for 4.6
The current SpaceXAI Grok 4.6 pages list a 500,000-token context window, text and image input, text output, function calling, structured outputs and reasoning. The documentation also names grok-4.6 as the model ID and shows Responses API plus OpenAI-compatible examples. [S43][S44]
| Field | Grok 4.6 |
|---|---|
| Context | 500,000 tokens [S43][S44] |
| Input/output | Text + image → text [S43][S44] |
| Base pricing | $2/$6; $0.50 cached input [S43] |
| Higher context | Higher input rates above 200K [S43] |
| Model ID | grok-4.6 [S44] |
| Reasoning | Low, medium, high, or xhigh [S44] |
These are product-level capabilities, not a promise that every gateway exposes every tool or that a Mac client will behave identically across providers.
The API model ID is another easy place to hallucinate
The current SpaceXAI documentation explicitly uses:
model="grok-4.6"
It also shows the model through the Responses API and an OpenAI-compatible base URL. [S43][S44] For developers, the robust migration process is still:
- enumerate models in the actual account;
- capture the exact 4.6 identifier, aliases and region;
- run a minimal text request;
- test structured output and function calling explicitly; [S15][S16]
- compare cost through request-level cost tracking; [S12]
- run the same agent benchmark against 4.5 before switching production traffic.
This is slower than copying a model ID from social media, but it prevents an avoidable production failure.
What Grok 4.5 tells us about the direction — not the 4.6 spec sheet
SpaceXAI’s 4.5 announcement presented manufacturer-reported results including 62.0% on DeepSWE 1.0, 53% on DeepSWE 1.1, 83.3% on Terminal-Bench 2.1 and 64.7% on SWE Bench Pro. [S07] The company also claimed high serving throughput and substantially lower token use on one software-engineering evaluation. Those are useful reference points, but they are vendor material and should not be blended with an independent 4.6 test.
Artificial Analysis independently measured 4.5 at 54 on its Intelligence Index and showed a large gain over Grok 4.3. [S26][S28] The post-launch report of another five-point move therefore makes 4.6 worth testing, especially for tasks that combine coding, tools and long sequences of decisions. [S01]
Who should test Grok 4.6 now?
Agent builders
The reported price/performance combination is most interesting for systems that repeatedly send large tool schemas, code context and task state. [S01][S02] Measure cost per completed task, not cost per million tokens.
Teams already on Grok 4.5
You have the cleanest A/B path: keep the harness, tools, prompts and evaluation set fixed, then swap only the model. Record success rate, human interventions, latency, token use and total cost.
Casual users
If the goal is simply “use the smartest chatbot,” one aggregate index is not enough evidence. The practical winner depends on your tasks, latency tolerance, product features and subscription constraints.
What remains open
The first-party docs settle the main product fields, but not every evaluation question. Still open or only partly comparable are:
- the final parameter count;
- a directly reproducible independent 4.6 leaderboard score;
- the full methodology behind the reported five-point improvement;
- regional pricing details for long contexts and third-party gateways;
- tool-calling, caching and long-agent reliability on a concrete Mac or provider setup.
These are not evidence against Grok 4.6. They are the boundaries of the current evidence and should be tested with a fixed workload before production migration.
Conclusion: Grok 4.6 is real, officially documented and available since August 12, 2026. First-party docs confirm its model ID, 500K context, modalities, tools and base pricing; the reported five-point improvement remains a qualified external claim. [S01][S43][S44] For Mac and agent workflows, run an A/B test of cost, latency, tool failures and task success instead of overfocusing on the unresolved parameter count.
Frequently Asked Questions
When was Grok 4.6 released?
Grok 4.6 launched on August 12, 2026. Multiple current reports place the launch on that date. [S01][S02]
How does Grok 4.6 compare with Grok 4.5?
MarketWatch reports a five-point lead over Grok 4.5 on the Artificial Analysis Intelligence Index. [S01] The exact current 4.6 score was not directly reproducible from the accessible leaderboard snapshot.
How much does Grok 4.6 cost via API?
SpaceXAI's official model page lists $2 per million input tokens and $6 per million output tokens. Cached input is $0.50 per million; higher input rates apply to prompts above 200,000 tokens. [S43][S44]
What is the Grok 4.6 context window?
SpaceXAI's official Grok 4.6 documentation lists a 500,000-token context window. That is the model maximum; provider limits, pricing tiers and latency still constrain practical use. [S43][S44]
Is Grok 4.6 a 1.5T or 2T model?
The figures conflict: older pre-launch reports say 2T, later Musk-attributed reporting says 1.5T. [S30][S32][S34] No accessible final model card existed at the cut-off, so the parameter count belongs in a "still to confirm" column.
What is the official Grok 4.6 API model ID?
The official model identifier is `grok-4.6`. Current documentation shows Responses API and OpenAI-compatible examples; still check available regions and limits in your own account. [S43][S44]
Transparency
Sources and review basis
These primary and reference sources form the basis of the technical assessment. Vendor claims and external benchmarks are identified as such in the article.
- marketwatch.com story / spacexs-stock-is-getting-a-grok-fueled-boost-e99497b4
- barrons.com articles / spacex-stock-grok-ai-anthropic-openai-musk-11939184
- axios.com newsletters / axios-am-bd5f21b3-41b2-48fc-829c-4be50402e91d
- investors.com news / spacex-stock-grok-openai-anthropic-elon-musk
- businessinsider.com spacex-cursor-acquisition-partnership-grok-bot-colossus-2026-8
- uol.com.br 13 / a-semana-em-que-zuckerberg-e-musk-sairam-do-z4-da-ia-e-botaram-medo-nos-lideres.ghtm
- x.ai news / grok-4-5
- docs.x.ai models / grok-4.5
- docs.x.ai developers / pricing
- docs.x.ai developers / release-notes
- docs.x.ai inference / models
- docs.x.ai developers / cost-tracking
- docs.x.ai developers / models
- docs.x.ai text / generate-text
- docs.x.ai text / structured-outputs
- docs.x.ai tools / function-calling
- docs.x.ai tools / remote-mcp
- docs.x.ai faq / security
- docs.x.ai advanced-api-usage / context-compaction
- docs.x.ai advanced-api-usage / priority-processing
- x.ai news
- x.ai news / grok-github-copilot
- x.ai news / grok-automations
- x.ai news / grok-build-mode
- apps.apple.com grok-ai-assistant / id6670324846
- artificialanalysis.ai models / grok-4-5
- artificialanalysis.ai articles / grok-4-5-brings-spacexai-to-the-the-intelligence-frontier
- artificialanalysis.ai comparisons / grok-4-5-vs-grok-4-3
- artificialanalysis.ai changelog
- pro.edgex.exchange article / grok-4-6-aug-7-grok-4-7-follows
- economictimes.indiatimes.com articleshow / 132693365.cms
- ithome.com 981 / 947.htm
- computerbase.de apps / start-von-grok-4-6-in-kuerze-xai-will-grok-zur-app-plattform-ausbauen.98525
- rb.ru news / ilon-mask-anonsiroval-srazu-dva-krupnyh-obnovleniya-nejroseti-grok-pervuyu-model-predstavyat-uzhe-7-avgusta
- testingcatalog.com spacexai-develops-deployable-applications-for-grok-build
- bighatgroup.com blog / xai-weekly-2026-07-29
- nextbigfuture.com 07 / spacexai-will-release-grok-4-6-in-2-weeks-and-grok-4-7-in-4-weeks.html
- aitoolsreview.co.uk insights / grok-4-6-grok-4-7-release-date
- nodemini.com blog / 2026-grok-4-6-release-date.html
- kie.ai blog / what-is-grok-4-6
- openrouter.ai models
- help.openai.com articles / 12627856-publishers-and-developers-faq
- docs.x.ai models / grok-4.6
- docs.x.ai developers / grok-4-6