Fetching from the wire…
01
02
03
04
05
06
07
08
Source-backed findings, relationship evidence, citations, and briefing history from the public MindPattern archive.
Olmo Hybrid 7B achieves 2x data efficiency over Olmo 3 on MMLU.
Source findingpaddo.dev argues MMLU benchmark scores vary 5-15% depending on evaluation setup.
Source findingGPT-OSS exhibits performative chain-of-thought on MMLU
Source findingDeepSeek-R1 exhibits performative chain-of-thought on MMLU achieving 80% token reduction
Source findingOlmo Hybrid 7B achieves 2x data efficiency over Olmo 3 on MMLU.
Source findingpaddo.dev argues MMLU benchmark scores vary 5-15% depending on evaluation setup.
Source findingGPT-OSS exhibits performative chain-of-thought on MMLU
Source finding