MIT Study: 41 AI Models Tested on 11,000 Real Tasks — Most Barely Meet Minimum Standards, Not Replacing Humans
MIT / Fortune·medium signal
A MIT study tested 41 AI models on over 11,000 real workplace tasks and found most AI outputs score well below 'superior' quality, barely meeting minimum sufficiency on complex work. The research estimates AI could handle 80-95% of text-based tasks at minimally sufficient level by 2029, but currently struggles with nuanced judgment and multi-step reasoning. The r/ChatGPT discussion (188 upvotes, 72 comments) frames this as the 'good enough problem' — AI that passes but doesn't impress.