Leading US AI lab maintains a 3-9-month capability lead
We imagine the others to be 3-9 months behind OpenBrain (page 4, Late 2025). By Early 2026, several competing AIs match or exceed Agent-0 (page 7). The 3-9 month gap is the Late 2025 state; near-parity emerges by Early 2026.
Late-2025 public competition was closer than the predicted sustained three-to-nine-month lead. Undisclosed internal advantages remain unresolved.
At a glance
- Assessment: Behind
- Confidence in assessment: 85%
- Outcome: unresolved
- Timing: Original window passed; outcome unresolved
- Evidence: proxy
- Predicted timing: Late 2025
- Primary source: ai-2027.com, Late 2025; AI Futures grading (Feb 2026)
What AI 2027 Predicted
The leading US lab maintains a three-to-nine-month advantage over competitors in late 2025. Evidence of a smaller lead therefore contradicts this particular prediction, even if it indicates faster competitive progress overall. AI 2027
How We Track This
Compare contemporaneous capabilities on common evaluations, rather than equating release frequency with capability. Distinguish observable public competition from hidden internal work.
Current Evidence
Epoch’s ECI table shows Gemini 3 Pro taking the public frontier on November 18, 2025, then GPT-5.2 Pro surpassing its point score 23 days later. Their year-end scores, 153.06 and 155.07 respectively, have overlapping 90% intervals. This supports close public competition rather than a clearly demonstrated sustained lead of several months. Epoch ECI table
The scenario’s authors independently judged the 2025 lead closer to zero to two months in their self-grading. That is the authors’ assessment, not an independent controlled measurement. AI Futures self-grading
Counterevidence & Limitations
An aggregate public index can obscure specialized strengths. Overlapping intervals do not prove exact equality, and public releases do not fully reveal internal capability. Those limits prevent a definitive claim about every lab’s private advantage.
Behind refers to the forecast advantage failing to appear clearly in public evidence. The previous interpretation treated narrower competition as confirmation, reversing the direction of the original claim. This correction preserves the evidence of convergence while assessing it against the stated prediction.
What Would Change Our Assessment
- Strengthen the original claim: Dated, credible internal evidence establishes a sustained three-to-nine-month lead in late 2025.
- Strengthen the shortfall assessment: Multiple comparable contemporary evaluations show persistent close competition across capability dimensions.
- Clarify: Document which capabilities and public or internal frontier the claimed advantage describes.
Update History
| Date | Update |
|---|---|
| 2026-09-07 | Reconciled close public competition with the direction of the original prediction. The dated ECI comparison supports a smaller public lead; this contradicts the forecast advantage, while internal gaps remain unresolved. |
| 2026-09-06 | Assessment revised from confirmed (0.90) to emerging (0.65). Release clustering and later competition do not resolve the original late-2025 three-to-nine-month capability lead. |
| 2026-06-22 | xAI made Grok 4.3 available on Amazon Bedrock and Anthropic launched Fable 5. These releases reinforce the confirmed assessment that top US labs remain clustered near the frontier. |
| 2026-04-13 | Anthropic surpassed OpenAI in revenue ($30B vs $24B ARR, April 2026). Anthropic achieved this while spending roughly 4x less on training (SaaStr, The AI Corner). The competitive gap between top US labs has fully closed — arguably inverted — with Anthropic, OpenAI, and Google trading frontier position every few weeks. Prediction firmly reinforced. |
| 2026-03 | Frontier race between OpenAI, Anthropic, and Google is extremely tight. The corrected 0-2 month gap is clearly borne out by alternating benchmark leads. |
| 2026-02 | AI Futures grading explicitly conceded: “The race appears to be closer than we predicted, more like a 0–2 month lead between the top US AGI companies.” The scenario’s 3–9 month OpenBrain lead did not materialize. |
| 2026-01 | AI Futures Project authors grade the inter-lab gap as 0-2 months, tighter than the 3-9 months predicted in AI 2027. AI Futures Project’s clarification post confirmed: “The race appears to be closer than we predicted, more like a 0-2 month lead between the top US AGI companies.” |
| 2025-12 | GPT-5.2 released three weeks after Gemini 3 triggered an internal “Code Red” at OpenAI. The reactive cadence — Gemini 3 (Nov 18) → GPT-5.2 (Dec 11) — illustrates a gap of weeks, not months, between top US labs. |
| 2025-11 | Gemini 3 and GPT-5.1-Codex-Max release on November 18; Claude Opus 4.5 releases on November 24 — six days later. Labs are trading benchmark leadership on a weekly basis. The AI 2027 prediction of 0-2 month gap between top labs appears confirmed. |
| 2025-08 | GPT-5 (August 7), Gemini 2.5 Deep Think availability expansion (August 1), and Claude Opus 4.1 (August 5) all release within days of each other. Multiple frontier-class releases in the same week from competing labs. |
| 2025-05 | Claude Opus 4 (May 22), Google Gemini 2.5 Flash GA (May 20), and OpenAI’s continued iteration: three major labs releasing competitive models within days of each other. The gap between top US labs is compressing. |
| 2025-04 | Meta Llama 4 (April 5, Apache 2.0) and Alibaba Qwen3 (April 28) both produce open-weight models competitive with some closed frontier models. The competitive surface for frontier capabilities is widening beyond the top three US labs. |