Pular para o conteúdo
← Back to Skalablog

Published article

What is the Grok 5 release date?

GrokAnthropicOpenAI

The Grok 5 release date is the most repeated and least supported number in AI right now. xAI has publicly targeted the model before the end of 2026, yet no architecture, context window, benchmark preview or training-compute figure has been published, so the honest answer separates a stated schedule from a measured capability.

What xAI has actually confirmed about Grok 5

The Grok 5 release date is a stated target for before the end of 2026, repeated by Elon Musk and tied to a training corpus claim, not a dated launch commitment. Everything beyond those two statements, including parameter count and any AGI probability, remains unverified as of September 2026.

xAI is Musk's AI company and the developer of the Grok assistant, distributed through X and a standalone app. On the company's own terms, Grok 5 is a planned model with a rough timeline. There is no public architecture document, no context-window specification, no benchmark preview and no disclosed training-compute figure.

The Grok 4.7 slip that reframed Grok 5 release date talk

Grok 4.7 was expected on September 11, 2026, and did not ship; two days later the conversation moved to Grok 4.8, which reframed every Grok 5 release date claim that followed. The sequence is short and matters more than any single promise because it shows how quickly the plan changed.

The published road map went through several revisions: Musk first described Grok 4.7 in late July 2026 as a 2.1 trillion parameter model arriving a few weeks after Grok 4.6; Grok 4.6 launched on August 12, 2026; the original late-August window closed with nothing shipped; September 1 brought a ten-day estimate; September 11 needed more time; and by September 13 the discussion had moved on to Grok 4.8.

Those dates and parameter figures come from the video's account of Musk's public statements rather than from an xAI specification sheet. Treat the 2.1 trillion figure as a claim made in public posts, not as a published model card.

One detail from that stretch is worth keeping. On the way out, Grok 4.7 was described as roughly on par with Anthropic Opus 5 rather than ahead of it. A model teased for six weeks as a major leap that lands level with an existing competitor is not a headline release, whatever the parameter count says.

What could the Colossus II GPU build-out deliver?

Colossus II is xAI's Blackwell-generation expansion in Memphis, and it is the part of the Grok 5 story that independent analysts can inspect rather than take on trust. The reported hardware matters to readers because training compute constrains what any model can plausibly learn.

xAI offers no published core-model card for Grok 5. Nvidia supplies the GB200 and GB300 accelerators discussed in the video as the Colossus II base. Reported plans describe roughly 110,000 GB200 GPUs delivering about 278 MW, plus a planned second tranche of about 330,000 GB300 chips, together pushing toward 440,000 Blackwell GPUs with an IT load above 946 MW. Those are reported plans, not an xAI capacity statement.

How much compute and cost sit behind the model

Reported GPU hardware spending for Colossus II is around $18 billion, with total infrastructure possibly reaching $35-58 billion once buildings, cooling, battery storage and power systems are counted. Those are estimates quoted in the video, not figures from a filing.

The power side is specific. The video describes 720 Tesla Megapacks and roughly 2.8-3.3 GWh of battery storage, plus on-site natural gas turbines for peak demand, a build-out that has drawn a lawsuit from local environmental groups over unpermitted emissions. That claim should be read as reported, not adjudicated.

Two ranges are worth naming plainly because they are the largest single points of uncertainty in the whole story:

  • Colossus II GPUs alone: about 440,000 reported Blackwell-class chips, with roughly 555,000 across the campus when Colossus I is included.
  • Power draw: some 946 MW of IT load in Colossus II, approaching 2 GW for the operation as a whole.

What would the compute actually buy in reasoning and coding?

More compute should extend how long a chain of reasoning Grok 5 can hold before it starts losing track of variables, but it does not automatically produce better coding or autonomous agents. The distinction decides which expectations are reasonable.

On reasoning, the current lineage is already capable within its stated limits. The video cites Grok 4.6 handling long logic chains inside a 500,000-token context and Grok 3 scoring 93.3% on the AIME math contest. Those are the video's figures; treat them as reported rather than reproduced.

On software engineering the picture is less flattering. The video puts Grok 4.6 at 56.4% versus 58.8% for its closest contemporaneous rival on Stanford's Apex SWE benchmark, with a wider gap on terminal and shell tasks at 26% versus 34%. Stanford HAI publishes research that includes such evaluations. Closing a coding gap usually requires better training data and reinforcement learning on real software tasks, not just a larger parameter count.

An entire SpaceX archive is not AGI

Training on more than 25 years of SpaceX engineering records would widen Grok 5's knowledge, but knowledge coverage is not general intelligence and does not change how the model learns after deployment. The distance between those two things is the core of the AGI question.

Musk has linked Grok 5 to the entire SpaceX data corpus. The video reports this as a real claim from Musk that no document independently verifies. A large, private engineering archive could plausibly improve technical reasoning if it were used well, and that 'if' carries most of the weight.

A static pre-trained model does not update its weights after training, needs explicit prompts to move between tasks and has no persistent memory across sessions by default. Against common working definitions of general intelligence, those are mechanical gaps that parameter count does not close. A larger context window holds a longer conversation; it does not create a goal-setting agent.

The probability figure that circulates with this discussion deserves the same care. The video traces a 10% chance of AGI line to a single interview from November 2025, before Grok 4.5 had shipped, and notes that nothing since has confirmed it. Attribute it to that interview and date it, or leave it out.

Grok 5 release date versus the shipping competition

By the time Grok 5 plausibly arrives, OpenAI, Google and Anthropic will have moved, which is why a single-digit benchmark gap is the realistic range for any launch claim. The video frames the field as converging rather than separating.

The named comparison points in September 2026 are OpenAI, whose most advanced model the video describes as GPT-6 Astra rolling out from September 4, 2026, stepping above the GPT-5.6 family from July 2026; Google, whose flagship reasoning model Gemini 3.1 Pro had not been refreshed since February 2026 while four separate flash-tier models shipped from May; and Anthropic, with Opus 5 and Sonnet 5 in general release and a Frontier Tier expanding access.

The video's reading of the Stanford AI Index is that the top labs now sit within a few percentage points of each other on standard benchmarks rather than the double-digit gaps of two or three years ago. Take the strategic conclusion as inference: if a single archive plus a larger training run were sufficient for a category leap, the field would not have converged this tightly.

Reading the next Grok 5 announcement

Ask three questions of the next Grok 5 announcement: is the claim dated, is the figure sourced to a specification or a post, and does the claimed capability change the learning mechanism? The answers sort predictions from marketing.

  • Treat a published model card with measured benchmark configuration as different evidence from a spoken target date.
  • Treat parameter counts as claims unless a training report or model card states them.
  • Treat any benchmark result the same way you would treat a competitor's: check the workload, the version and who ran it.

The pattern worth watching is the pivot itself. Grok 4.7 gave way to Grok 4.8 within two days in September 2026, which suggests the road map is written close to the event.

FAQ

  • When is Grok 5 coming out? xAI has publicly targeted Grok 5 before the end of 2026 through repeated statements from Elon Musk, including an August 2026 earnings call. No launch date, model card or architecture document has been published as of September 2026. Treat the date as a stated target rather than a schedule.
  • Does Grok 5 have 6 trillion parameters? The six trillion parameter figure traces to a single interview in November 2025, before Grok 4.5 had shipped. Nothing published since has confirmed it, and xAI has not released a parameter count. It is best treated as a low-confidence claim.
  • Will Grok 5 be AGI? A larger pre-trained model can score well across more benchmarks simultaneously, which is not the same as general intelligence. Grok 5 would still lack post-deployment learning, persistent memory by default and independent goal setting. No public Grok 5 material addresses those mechanisms.
  • How many GPUs does Colossus II have? Reported plans describe roughly 440,000 Blackwell-generation GPUs across the Colossus II expansion, about 110,000 GB200 plus a planned 330,000 GB300 tranche, with roughly 555,000 across the wider campus including Colossus I. These are reported plans rather than an xAI capacity statement.
  • Will Grok 5 beat GPT-6 Astra or Gemini 3.1 Pro? The video's forecast is that Grok 5 lands competitively against whatever OpenAI, Google and Anthropic shipping rather than decisively ahead. The basis is that top-lab scores have converged, so single-digit margins are the realistic range for any launch gap.
  • Is Grok a single multimodal model? No. Grok's chat model calls out to separate systems for image generation and voice, and video is a third system, so it is modular rather than one network that natively processes pixels and audio. The video expects Grok 5 to remain modular rather than a unified multimodal network.

Turning dense AI roadmaps into a written explainer

Grok 5 coverage is a good example of a video that carries more value in its structure than in any single claim: a timeline, a set of confidence levels and a clear split between what is confirmed and what is inferred. That structure is what makes it worth reading rather than watching at 1.5x.

If you publish conversations or explanations like this, the same material can live as a written article. Paste a YouTube URL into Skalablog, let it transcribe the video and generate a draft article you can edit. Written pages are easier to search, cite and update than a video description.

That matters most for claims that age fast. A dated article can mark each figure with its source and its confidence level, so a reader can tell what changed. With Skalablog the transcription and first draft come from the video you already made.

Source video