Early user tests showed OpenAI's GPT-6 Astra handling browser, desktop and visual tasks directly from prompts. On Browser Use Benchmark v2, Astra medium scored 77.3% versus Opus 5's 50.5%, and Astra high scored 92.9% on WeirdML, matching Fable 5.1 max for the top score.

Separate demonstrations showed Astra rebuilding an apartment display and adding a telescope feed in about 20 minutes, and other users said it solved logic puzzles without code. Some users also said Astra consumed tokens quickly, with one Pro x20 subscriber saying a day of Astra Light use nearly exhausted the plan's weekly usage.

Sign in to suggest edits

Key sources

  1. SOURCE@sebastienbubeck“Anything you can do on a computer, Astra can do for you. Fast.”x.com
  2. SUPPORT@npew“Astra just obliterated our hardest benchmark 👀”x.com
  3. SUPPORT@rhyssullivan“20 minutes later it was done and worked great”x.com
  4. SUPPORT@teortaxestex“GPT-6 Astra is the best "vision" model I've seen”x.com
  5. SUPPORT@scaling01“it logically (no code) solved ALL of them”x.com
  6. SUPPORT@htihle“It sets a new individual high score on 6 of the 17 tasks”x.com
  7. SUPPORT@koltregaskes“basically out of my weekly usage on Pro x20 plan”x.com
  8. SOURCE@brendanfoody“Anything you can do on a computer, Astra can do for you. Fast.”x.com
Markdown