Fetching from the wire…
Source-backed findings, relationship evidence, citations, and briefing history from the public MindPattern archive.
Showing the first 40 findings. More graph evidence exists in the corpus.
V4 Flash 0731 achieved 61.4% accuracy on ARC-AGI-2 benchmark at $0.04 per task.
Source findingInkling scored 36.5% on ARC-AGI-2.
Source findingHourglass Reasoning improves ARC-AGI-2 best-of-5 accuracy by 14 points.
Source findingSymbolica Agentica achieved 85.28% on ARC-AGI-2
Source findingPoetiq achieved 54% on ARC-AGI-2 via iterative refinement
Source findingGemini 3.1 Pro achieved 77.1% on ARC-AGI-2 benchmark.
Source findingGemini 3.1 Pro scored 77.1% on ARC-AGI-2.
Source findingGemini 3.1 Pro achieved 77.1% on ARC-AGI-2 benchmark.
Source findingMeta^n beats prior self-improving agents on ARC-AGI-2 without memorizing skills.
Source findingV4 Flash 0731 achieved 61.4% accuracy on ARC-AGI-2 benchmark at $0.04 per task.
Source findingInkling scored 36.5% on ARC-AGI-2.
Source findingHourglass Reasoning improves ARC-AGI-2 best-of-5 accuracy by 14 points.
Source finding