Fetching from the wire…
Source-backed findings, relationship evidence, citations, and briefing history from the public MindPattern archive.
Showing the first 40 findings. More graph evidence exists in the corpus.
Specific Labs published Real-SWE, a benchmark using private production codebases.
Source findingGemini 3.8 Flash achieved 31.2% pass@1 on Real-SWE's benchmark.
Source findingGPT-6 Astra achieved 33.8% pass@1 on Real-SWE's private codebase benchmark.
Source findingFable 5.1 achieved 38.8% pass@1 on Real-SWE's private codebase benchmark tasks.
Source findingReal-SWE benchmarked GPT-6 Astra on private codebases, measuring 33.8% pass@1.
Source findingSpecific Labs published Real-SWE, a benchmark using private production codebases.
Source findingGemini 3.8 Flash achieved 31.2% pass@1 on Real-SWE's benchmark.
Source findingReal-SWE benchmarked GPT-6 Astra on private codebases, measuring 33.8% pass@1.
Source finding