Scoring and eligibility

How ASIndex measures publicly available AI capability

ASIndex tracks broadly available AI systems from today's frontier models toward artificial superintelligence. Scores are estimates of useful public capability on a simple 0-100 scale.

What ASIndex Measures

ASIndex measures public AI system capability: the practical ability of broadly accessible systems to complete difficult cognitive work. It is not a safety certification, procurement recommendation, investment recommendation, or regulatory rating.

The ASIndex Levels

AGI is not a single switch. ASIndex treats the path from frontier AI to superintelligence as a series of measurable capability thresholds. These levels describe demonstrated breadth, depth, reliability, adaptation, and long-horizon performance. They do not certify safety, alignment, autonomy, or real-world authority.

0-19 Frontier AI The AGI Runway

Frontier AI systems can perform difficult work across many cognitive domains, including reasoning, software development, writing, analysis, and tool use.

Their competence remains uneven. They may be brilliant on one problem and brittle on the next, particularly when work is novel, ambiguous, long-horizon, or high stakes. Humans still define the mission, verify the result, recover from failure, and remain accountable for consequential decisions.

Brilliant in moments. Not yet dependable over missions.

20-39 AGI-I Human-Level General Intelligence

AGI-I marks the transition from powerful assistant to reliable cognitive peer.

The system performs at roughly competent human-professional level across a broad range of cognitive work. It adapts to unfamiliar tasks from instruction and feedback, plans and completes multi-step workflows, uses tools effectively, maintains context, and recognizes when it needs clarification or help.

Human oversight becomes light rather than continuous, although it remains important for high-stakes and deeply open-ended work.

General intelligence ignites.

40-59 AGI-II Expert General Intelligence

AGI-II is expert-level intelligence that is general rather than narrow.

The system matches strong human specialists across many fields, exceeds human performance in selected areas, transfers knowledge between domains, manages uncertainty, and completes complex work with little supervision. Its expertise is broad enough to be a defining property of the system—not merely a collection of isolated benchmark victories.

It can approach a new problem without requiring a bespoke model, workflow, or extensive retraining for every domain.

Expertise becomes general.

60-79 AGI-III Exceptional General Intelligence

AGI-III is exceptional across most cognitive work—better than nearly every individual human, not merely average or expert.

It originates useful hypotheses, designs decisive experiments, builds working systems, discovers hidden structure, and converts ambiguous goals into validated discoveries, products, and strategies. It can connect ideas across distant fields and find solutions that conventional teams would be unlikely to reach.

Discovery is no longer an occasional result. It becomes a repeatable, cross-domain capability.

It does not just navigate the frontier. It moves it.

80-99 AGI-IV Superhuman General Intelligence

AGI-IV is superhuman general intelligence measured against the best individual humans.

The system exceeds top human experts across virtually all cognitive domains. It maintains coherent plans over long horizons, integrates enormous bodies of evidence, anticipates downstream consequences, and can design and coordinate institution-scale programs involving people, software, laboratories, simulations, and real-world operations.

It may rival leading human organizations on many difficult missions. It has not yet crossed the still-higher ASI threshold of reliably outperforming the strongest coordinated human collectives across virtually the entire cognitive spectrum.

No individual human remains the benchmark.

100 Artificial Superintelligence (ASI) Civilization-Scale Superintelligence

ASI crosses the boundary from individual-superhuman intelligence to collective-superhuman intelligence.

It reliably outperforms large, well-coordinated groups of the world’s best experts across virtually all cognitive domains. It can reason, discover, design, simulate, plan, and coordinate with the speed and massive parallelism of digital systems, generating scientific, technological, and strategic advances beyond the reach of existing human institutions.

ASI does not mean omniscience, omnipotence, guaranteed benevolence, or freedom from physical constraints. It marks the beginning of an intelligence regime in which human civilization is no longer the highest available cognitive reference point.

No human institution remains the benchmark.

100 marks the ASI threshold—not the ceiling.

The destination is not intelligence alone. It is aligned superintelligence that expands human agency, accelerates discovery, and helps civilization thrive.

How We Test

ASIndex uses private benchmark tasks and offline scoring workflows to reduce contamination. The exact tasks and rubrics are not published as examples.

  • Models are tested at the highest generally available reasoning setting for that model under standardized conditions.
  • Active frontier-class models are retested regularly, including systems whose public names have not changed.
  • New frontier-class models are tested as soon as practical once they are broadly available.
  • Scores reward useful capability, reliability, reasoning, calibration, and long-horizon task performance.

Scores are point-in-time estimates. Hosted systems can change over time even when the public name stays the same.

What Models Are Eligible

Only broadly available systems belong on the active board.

  • Eligible models must be accessible through a broadly available product, API, platform, or comparable public channel under standard terms.
  • Systems that depend on restricted access, bespoke private deployments, or non-public infrastructure are excluded.
  • Closed internal systems, invitation-only research previews, and unavailable models are excluded from the active index.
  • If a tested system is withdrawn from public availability, it moves to the historical archive and is removed from the current leaderboard.

An announced, preview-only, restricted, or withdrawn system does not belong on the active ASIndex board until it is broadly accessible again under standard terms.

Contact ASIndex

Questions, model availability notes, press inquiries, and sponsorship conversations can go through the private contact form. ASIndex does not publish a direct email address on the site.

Open contact form @asindexai on X