ALE: GPT-5.5 edges Claude Fable 5 by 2 points; most tasks still fail
The new Agents' Last Exam (ALE) benchmark from UC Berkeley's Center for Responsible, Decentralized Intelligence (RDI) crowns OpenAI's gpt-5-5 with a 24.0% pass rate, with Anthropic's just-released claude-fable-5 sitting in third place at 22.0%. The article