Reddit
Vals AI's RSI Index: Opus 5.5 beats the reference result on one of five autonomous AI R&D tasks
The Vals RSI Index, updated 9/21, scores models on five open-ended LLM R&D tasks: compression, LM training, parameter golf, harness engineering and post-training. Each runs in the model's own harness at max effort (Claude in Claude Code, GPT in Codex). On the log scale, 0.5 means matching a published, human or frontier reference and 0.6 means beating it substantially. The r/singularity post (117 upvotes) reports that Opus 5.5 passed the reference on one task and moved the extrapolated full-RSI date from August to July 2027. Commenters noted that RSI will first happen on internal models, which public benchmarks cannot see.
↳ Follow the thread