OpenAI’s New AI Model "GPT-6 Astra" Sweeps All 48 Stages of "Human Proof" Captcha Test
Eugenio Rodolfo Sanabria Reporter
| 2026-09-09 09:31:04
OpenAI's latest artificial intelligence model, "GPT-6 Astra," has successfully solved and passed all stages of a popular online Captcha puzzle game designed to prove human identity on the internet.
According to Sharif Shameem, a developer at OpenAI, GPT-6 Astra was recently tested on the 48-stage online puzzle game titled "I'm Not a Robot," created by developer Neal Agarwal. By successfully clearing every single phase of the test, the AI model practically earned a metaphorical "human certificate."
The "I'm Not a Robot" challenge goes far beyond standard verification prompts. It is a comprehensive 48-level puzzle spanning image classification, drawing perfect circles, keeping up with rhythm matching, playing chess, and engaging in multi-turn dialogues. The stages require recognizing distorted alphabet characters, differentiating cookies from dogs in ambiguous photographs, and pinpointing specific urban structures like traffic lights within complex visual grids.
Most of these challenges are actively deployed across the internet to block automated web-scraping bots. Because of their high difficulty level, even real human users frequently make mistakes or fail certain rounds. Astra's flawless sweep of the test highlights its dramatic leap in multidisciplinary capabilities, including spatiotemporal reasoning through advanced image recognition, deep context comprehension of complex instructions, and seamless computer-use proficiency.
As commercially available AI models systematically neutralize security tools long trusted to separate humans from bots, cybersecurity experts note that traditional anti-bot and CAPTCHA mechanisms face an urgent need for massive overhauls.
Meanwhile, prominent AI performance evaluation firm Artificial Analysis adjusted its Artificial Analysis Intelligence Index (AAII) methodology twice within a four-day window following Astra's launch, drawing significant industry attention. On its debut day, September 4, Astra scored 61 points under the previous evaluation framework (version 4.1), placing it in a joint 5th place behind leading competitors such as Anthropic’s "Claude Fable 5.1," which scored 66 points. However, after implementing the updated v4.3 framework—which incorporates forward-looking metrics anticipating upcoming evaluation standards—Astra's adjusted scoring placed it tied for 1st place alongside Fable 5.1.
Industry analysts point out that these benchmark updates reflect the rapid velocity of modern AI development, prompting evaluation platforms to continuously evolve their testing criteria to match state-of-the-art frontier models.
WEEKLY HOT
- 1Pohang Arts High School Students Showcase Artistic Vision and Creativity at Gyeongju Expo Grand Park Exhibition
- 2Busan Gripped by "Busan Fever" as Taiwan’s Tourism Wave Surges
- 3KTO Launches Public-Private Consultative Body to Revitalize Inbound Car Ferry Tourism at Incheon Port
- 4KOICA Unveils Year 3 of its Branding Campaign “KOICA”: A Magical Journey of International Development Cooperation
- 5Japanese Culinary Students Experience the Depth of Jeonbuk Beyond Recipes
- 6From the Galapagos to Santorini: Capturing the Unique Charms of Global Islands at the 2026 Yeosu World Island Expo