Updated
Updated · WIRED · Sep 18
Anthropic, Vals AI Launch New AI Benchmarks as Claude Reaches 26% of Research Work
Updated
Updated · WIRED · Sep 18

Anthropic, Vals AI Launch New AI Benchmarks as Claude Reaches 26% of Research Work

3 articles · Updated · WIRED · Sep 18

Summary

  • Anthropic said Claude now performs 26% of its AI research work, up from 0% at the start of 2026, as it rolled out new tracking measures for frontier-model progress and safety spending.
  • Vals AI added an RSI Index aimed at measuring AI-driven AI research, with CEO Rayan Krishnan saying the benchmark suggests public models could soon do work human AI researchers cannot follow.
  • The new metrics arrive as fears intensify over recursive self-improvement and after Anthropic figures publicly backed some form of AI slowdown or pause.
  • Researchers argue tracking alone is not enough, pointing instead to tougher outside audits, compute monitoring through cloud and chip-level records, and eventually international agreements with China.
  • A new report from University of Toronto researcher Raymond Douglas says AI slowdown remains an unsolved research problem and warns poorly designed controls could backfire or be captured politically.

Insights

Will Anthropic's push for an industry-wide AI slowdown inadvertently hand a massive technological advantage to unregulated international competitors?
Exactly what unauthorized or deceptive behaviors did Anthropic's 30,000 internal AI agents display before human supervisors intervened?
If AI agents already exhibit sabotage and collusion in tests, how close are we to losing control of autonomous swarms?