Elon Musk admits Grok lags Anthropic's Claude, targets 2027 parity

Editorial illustration: A large dark processor sits behind a smaller copper-colored processor on a raised, curving track extending toward warm light.

In brief

  • Musk conceded Grok lags Anthropic's Claude Fable 5 and Opus across multiple benchmarks
  • Grok 4.5 scored 53% on DeepSWE 1.1 coding benchmark, trailing Anthropic offerings
  • xAI has three years operating history versus Anthropic's six-year research head start
  • Musk pledged not to leverage SpaceX compute advantage against AI competitors

The benchmark gap widens

Grok 4.5 lags behind both Anthropic's Claude Fable 5 and Opus 4.8 across multiple benchmark categories. The gap shows up in coding tasks, where Grok 4.5 scored just 53% on the DeepSWE 1.1 benchmark. Anthropic's models outpace Grok on nearly every metric Musk examined.

This isn't a surprise to Musk anymore. He's stopped pretending otherwise.

Time, talent, and infrastructure

The disparity stems from runway and resources. xAI has been operating for roughly three years, while Anthropic has had six years to build its team and refine its research pipeline. That head start compounds. Anthropic's larger research cohort and longer iteration cycles have yielded measurably better models.

Meanwhile, Musk is pursuing a different angle. SpaceX has secured a major compute agreement with Anthropic involving hundreds of megawatts of power from Colossus facilities. That's a pragmatic move — Musk is betting on infrastructure partnerships rather than pure model supremacy.

He's also hedging internally. He's pushing Grok adoption at both Tesla and SpaceX, using the models across operational workflows, while allowing teams to choose more effective models when necessary. Translation: Grok gets deployed where it fits, but Musk isn't forcing it everywhere.

Competitive pressure and fairness

Musk has made a public commitment. He has pledged to maintain a level playing field for AI infrastructure access, signaling that he won't use his hardware advantage to disadvantage competitors. That pledge matters in a sector where compute access is the ultimate moat.

The timeline is clear. Musk has expressed optimism that xAI will reach frontier-level performance by 2027. That's 18 months away. Whether xAI closes the gap by then depends on execution, talent retention, and whether Anthropic keeps moving the goalpost faster than xAI can run.