Fetching from the wire…
Source-backed findings, relationship evidence, citations, and briefing history from the public MindPattern archive.
Showing the first 40 findings. More graph evidence exists in the corpus.
GPT-5.1 was evaluated against GraphWalker for model-based test path generation.
Source findingOpenHands was used with GPT-5.1 as the baseline for comparison.
Source findingGemini 3.1 scored approximately 400 rating points higher than GPT-5.1 in chess benchmarks.
Source findingSimon Willison discussed AI coding capabilities as of GPT-5.1 release.
Source findingOpenAI published post-mortem on GPT-5.1 goblin-mention bug in reward signal
Source findingClaude Sonnet 4.5 and GPT-5.1 were both leading models at different points in November 2025.
Source findingOpenHands was used with GPT-5.1 as the baseline for comparison.
Source findingGemini 3.1 scored approximately 400 rating points higher than GPT-5.1 in chess benchmarks.
Source findingSimon Willison discussed AI coding capabilities as of GPT-5.1 release.
Source findingOpenAI published post-mortem on GPT-5.1 goblin-mention bug in reward signal
Source findingClaude Sonnet 4.5 and GPT-5.1 were both leading models at different points in November 2025.
Source finding