GPT-6 Astra Leads Coding and Math Benchmarks, Shows Strong Computer Use
Original titleGPT-6 Astra, Looped Transformers, and Hidden Reasoning
AISummary
OpenAI's GPT-6 Astra scores 99.9% on ARC-AGI-3, versus 7.8% for GPT-5.6 Sol, and leads Raschka's coding and math tests. Its strongest showing is in graphics and computer-use tasks, such as redrawing an image in a browser-based Paint app. The author notes that Artificial Analysis shows Astra at the frontier but not pulling far ahead on its Coding Agent Index.
Source: Ahead of AI (Sebastian Raschka) · magazine.sebastianraschka.com