← Về trang chính1 tin
Mô hình GPT-6 Astra vừa ghi nhận kết quả 14% trong bài benchmark MazeBench, cho thấy hiệu suất vượt trội gấp 7 lần so với đối thủ Claude Fable 5.1.
Đăng nhập để góp ý, chỉnh sửa
Nguồn chính
- SOURCE@htihle“Astra spent 60+ hours in this 3D open world spatial reasoning eval”x.com
- SUPPORT@rohanpaul_ai“reduced a projected 3B-token trajectory to about 350M tokens”x.com
- SOURCE@kimmonismus“Astra Pro scores 86.5%, almost tied with Claude Fable 5.1 at 86.6%”x.com
- SUPPORT@reach_vb“Astra is SoTA on MazeBench by a huge margin”x.com
- SUPPORT@andrewcurran_“massive jump with Astra”x.com