---
format: "aidr-story-markdown/v1"
id: "5356e63d435ad37e0456813c10b3bf1b7ddd20d613034fedce3b81db5a963a40"
canonical_url: "https://aidr.today/5356e63d?lang=en"
title: "GPT-6 Astra Seizes No. 1 on Code Arena WebDev With 35 Point Lead Over Claude"
lang: "en"
requested_lang: "en"
available_langs: ["en","vi"]
translation_fallback: null
fallback_fields: []
published_at: "2026-09-05T20:40:21.000Z"
category: "Models"
topics: ["openai","gpt","anthropic","claude","agent","benchmark","reasoning","fable"]
source_urls: ["https://x.com/arena/status/2096290434700247250","https://x.com/arena/status/2096292448960217449","https://x.com/wallstengine/status/2096310044291916063","https://x.com/saxena_puru/status/2096243535100129311","https://x.com/spicey_lemonade/status/2096283173059723701","https://x.com/j_dekoninck/status/2096276558600093969","https://x.com/firstadopter/status/2096283471714845099","https://x.com/jukan05/status/2096196101602996547"]
summary: "OpenAI's newest artificial intelligence model achieved a score of 1,797 on the Code Arena WebDev benchmark, taking the top spot from Anthropic's Claude Fable 5.1. The GPT-6 Astra (Max) model matches competitive pricing at $40 per million tokens and beats OpenAI's previous GPT-5.6 Sol entry by 180 points. In PINNACLE enterprise testing, the model completed 279 of 280 multi-step jobs at a cost of $1.51 per successful task, which is 39% less than Claude's $2.46. At its maximum reasoning setting, Astra fabricated no answers on questions that could not be answered from provided information. The update moves AI capability toward direct computer control of browsers and terminals to execute multi-step workflows, increasing token volume and compute demand per job. Initial access is limited to cybersecurity customers, with API and AWS integrations following. This agentic shift puts structural pressure on seat-based software pricing by automating workflows that previously required individual human user seats, shifting value toward compute infrastructure and data governance platforms."
---

# GPT\-6 Astra Seizes No\. 1 on Code Arena WebDev With 35 Point Lead Over Claude

> [Open the canonical story](<https://aidr.today/5356e63d?lang=en>)

**Published:** 2026-09-05T20:40:21.000Z
**Category:** Models
**Topics:** openai, gpt, anthropic, claude, agent, benchmark, reasoning, fable

## Summary

OpenAI's newest artificial intelligence model achieved a score of 1,797 on the Code Arena WebDev benchmark, taking the top spot from Anthropic's Claude Fable 5\.1\. The GPT\-6 Astra \(Max\) model matches competitive pricing at $40 per million tokens and beats OpenAI's previous GPT\-5\.6 Sol entry by 180 points\. In PINNACLE enterprise testing, the model completed 279 of 280 multi\-step jobs at a cost of $1\.51 per successful task, which is 39% less than Claude's $2\.46\. At its maximum reasoning setting, Astra fabricated no answers on questions that could not be answered from provided information\. The update moves AI capability toward direct computer control of browsers and terminals to execute multi\-step workflows, increasing token volume and compute demand per job\. Initial access is limited to cybersecurity customers, with API and AWS integrations following\. This agentic shift puts structural pressure on seat\-based software pricing by automating workflows that previously required individual human user seats, shifting value toward compute infrastructure and data governance platforms\.

## Sources

- [Story source](<https://x.com/arena/status/2096290434700247250>)
- [Supporting source](<https://x.com/arena/status/2096292448960217449>)
- [Supporting source](<https://x.com/wallstengine/status/2096310044291916063>)
- [Supporting source](<https://x.com/saxena_puru/status/2096243535100129311>)
- [Supporting source](<https://x.com/spicey_lemonade/status/2096283173059723701>)
- [Story source](<https://x.com/j_dekoninck/status/2096276558600093969>)
- [Supporting source](<https://x.com/firstadopter/status/2096283471714845099>)
- [Supporting source](<https://x.com/jukan05/status/2096196101602996547>)

