Skip to content

PreviewAll content, scores and forecasts here are illustrative sample data — not reporting, and not measurements.What this means

Agitology

AGI Index dimension · weight 9%

Autonomy

Useful operation without a human in the loop.

How long a system can run unsupervised on a real task before human intervention is required to keep it correct, safe or on-budget. Measured in wall-clock time and in consequential decisions taken without review.

The argument

Evidence and counter-evidence

Both sides are published at equal weight. A framework that only records what raises a score is not measuring anything.

Raises the score

  • Unsupervised operating windows have lengthened from minutes to hours on well-specified engineering work.
  • Self-monitoring — recognising that a run has gone off the rails and stopping — is beginning to appear.

Holds it down

  • Deployment practice remains overwhelmingly supervised, which is itself evidence about trust.
  • Failure modes are silent: systems more often produce plausible wrong work than halt, which raises the cost of removing the human.
  • Autonomy is measured almost entirely in domains with cheap, fast feedback. It is largely untested where errors are expensive.

What the score reads from

  • Unsupervised operating window

    Hours on constrained tasks; minutes on open-ended ones.

  • Intervention rate

    Falling steadily, still far from the threshold that would justify unsupervised deployment.