Reddit
Unverified: r/singularity Post Claims GPT-5.6 Sol Reaches the ZeroBench Human Baseline at pass@5 Without Tools
A 138-upvote r/singularity post claims GPT-5.6 Sol hits the human baseline on ZeroBench — the deliberately 'impossible' visual benchmark for multimodal models — at pass@5 with no tools, and the poster is careful to distinguish pass@5 (correct on at least one of five attempts) from pass^5 and from an averaged pseudo-pass@1. Flagging this as unconfirmed: the claim does not appear in OpenAI's published GPT-5.6 materials or on the ZeroBench results table, where contemporary models score between 2 and 19 out of 100 on the main questions. If it holds it is a meaningful multimodal milestone, but the pass@5 framing does a lot of work and nobody should cite the number until a primary source publishes it.
↳ Follow the thread