Leading US AI lab maintains a 3-9-month capability lead

Last updated
We imagine the others to be 3-9 months behind OpenBrain (page 4, Late 2025). By Early 2026, several competing AIs match or exceed Agent-0 (page 7). The 3-9 month gap is the Late 2025 state; near-parity emerges by Early 2026.

Late-2025 public competition was closer than the predicted sustained three-to-nine-month lead. Undisclosed internal advantages remain unresolved.

At a glance

  • Assessment: Behind
  • Confidence in assessment: 85%
  • Outcome: unresolved
  • Timing: Original window passed; outcome unresolved
  • Evidence: proxy
  • Predicted timing: Late 2025
  • Primary source: ai-2027.com, Late 2025; AI Futures grading (Feb 2026)

What AI 2027 Predicted

The leading US lab maintains a three-to-nine-month advantage over competitors in late 2025. Evidence of a smaller lead therefore contradicts this particular prediction, even if it indicates faster competitive progress overall. AI 2027

How We Track This

Compare contemporaneous capabilities on common evaluations, rather than equating release frequency with capability. Distinguish observable public competition from hidden internal work.

Current Evidence

Epoch’s ECI table shows Gemini 3 Pro taking the public frontier on November 18, 2025, then GPT-5.2 Pro surpassing its point score 23 days later. Their year-end scores, 153.06 and 155.07 respectively, have overlapping 90% intervals. This supports close public competition rather than a clearly demonstrated sustained lead of several months. Epoch ECI table

The scenario’s authors independently judged the 2025 lead closer to zero to two months in their self-grading. That is the authors’ assessment, not an independent controlled measurement. AI Futures self-grading

Counterevidence & Limitations

An aggregate public index can obscure specialized strengths. Overlapping intervals do not prove exact equality, and public releases do not fully reveal internal capability. Those limits prevent a definitive claim about every lab’s private advantage.

Behind refers to the forecast advantage failing to appear clearly in public evidence. The previous interpretation treated narrower competition as confirmation, reversing the direction of the original claim. This correction preserves the evidence of convergence while assessing it against the stated prediction.

What Would Change Our Assessment

  • Strengthen the original claim: Dated, credible internal evidence establishes a sustained three-to-nine-month lead in late 2025.
  • Strengthen the shortfall assessment: Multiple comparable contemporary evaluations show persistent close competition across capability dimensions.
  • Clarify: Document which capabilities and public or internal frontier the claimed advantage describes.

Update History

DateUpdate
2026-09-07Reconciled close public competition with the direction of the original prediction. The dated ECI comparison supports a smaller public lead; this contradicts the forecast advantage, while internal gaps remain unresolved.
2026-09-06Assessment revised from confirmed (0.90) to emerging (0.65). Release clustering and later competition do not resolve the original late-2025 three-to-nine-month capability lead.
2026-06-22xAI made Grok 4.3 available on Amazon Bedrock and Anthropic launched Fable 5. These releases reinforce the confirmed assessment that top US labs remain clustered near the frontier.
2026-04-13Anthropic surpassed OpenAI in revenue ($30B vs $24B ARR, April 2026). Anthropic achieved this while spending roughly 4x less on training (SaaStr, The AI Corner). The competitive gap between top US labs has fully closed — arguably inverted — with Anthropic, OpenAI, and Google trading frontier position every few weeks. Prediction firmly reinforced.
2026-03Frontier race between OpenAI, Anthropic, and Google is extremely tight. The corrected 0-2 month gap is clearly borne out by alternating benchmark leads.
2026-02AI Futures grading explicitly conceded: “The race appears to be closer than we predicted, more like a 0–2 month lead between the top US AGI companies.” The scenario’s 3–9 month OpenBrain lead did not materialize.
2026-01AI Futures Project authors grade the inter-lab gap as 0-2 months, tighter than the 3-9 months predicted in AI 2027. AI Futures Project’s clarification post confirmed: “The race appears to be closer than we predicted, more like a 0-2 month lead between the top US AGI companies.”
2025-12GPT-5.2 released three weeks after Gemini 3 triggered an internal “Code Red” at OpenAI. The reactive cadence — Gemini 3 (Nov 18) → GPT-5.2 (Dec 11) — illustrates a gap of weeks, not months, between top US labs.
2025-11Gemini 3 and GPT-5.1-Codex-Max release on November 18; Claude Opus 4.5 releases on November 24 — six days later. Labs are trading benchmark leadership on a weekly basis. The AI 2027 prediction of 0-2 month gap between top labs appears confirmed.
2025-08GPT-5 (August 7), Gemini 2.5 Deep Think availability expansion (August 1), and Claude Opus 4.1 (August 5) all release within days of each other. Multiple frontier-class releases in the same week from competing labs.
2025-05Claude Opus 4 (May 22), Google Gemini 2.5 Flash GA (May 20), and OpenAI’s continued iteration: three major labs releasing competitive models within days of each other. The gap between top US labs is compressing.
2025-04Meta Llama 4 (April 5, Apache 2.0) and Alibaba Qwen3 (April 28) both produce open-weight models competitive with some closed frontier models. The competitive surface for frontier capabilities is widening beyond the top three US labs.