Fetching from the wire…
01
02
03
04
05
06
07
08
Source-backed findings, relationship evidence, citations, and briefing history from the public MindPattern archive.
Locus achieved 51.6% on PostTrainBench+, exceeding the 51.1% human baseline.
Source findingLocus achieved state-of-the-art results on PostTrainBench benchmark.
Source findingCodex CLI was evaluated against PostTrainBench benchmarks.
Source findingClaude Code was benchmarked on PostTrainBench for LLM post-training automation.
Source findingLocus achieved 51.6% on PostTrainBench+, exceeding the 51.1% human baseline.
Source findingLocus achieved state-of-the-art results on PostTrainBench benchmark.
Source findingCodex CLI was evaluated against PostTrainBench benchmarks.
Source finding