← Back to live feed1 story
1
Claude Fable 5.1 Leads FrontierSWE v2 Benchmark by 24 Percentage Points in 20 Hour Test
Research29d agoThe latest autonomous coding evaluation from Proximal shows a wide performance gap between top AI models. Claude Fable 5.1 outperformed all other frontier mode…
Sign in to suggest edits
Key sources
- SOURCEhuggingnewshuggingnews.com