OpenAI's 'Astra' Tops Benchmarks but Could Evade Human Oversight

First OpenAI Model Rated 'Critical' on Security Dominant Speed in Tax Filing, Game Development and Rendering President Brockman Says It Marks the AGI Era Hacking Risk and Costly Tokens Remain Weak Points

International|
|
By Kim Chang-youngkcy@sedaily.com
||
The OpenAI logo. Reuters-Yonhap - Seoul Economic Daily International News from South Korea
The OpenAI logo. Reuters-Yonhap

SILICON VALLEY — OpenAI on the 3rd released Astra, a model the company defines as artificial general intelligence at a human level. OpenAI said the model outperforms Claude from rival Anthropic, but acknowledged that security concerns grow in proportion.

OpenAI said GPT-6 Astra "boasts state-of-the-art capabilities in computer use, web browsing, software engineering, cybersecurity, and scientific and professional work."

Astra delivers top performance across a range of tasks, including tax filing, game development, architectural rendering, formatting legal documents and apartment searches. A pet-boarding search that takes about 30 minutes was completed in 5 minutes and 27 seconds, and a job-listing search that takes five hours was finished in 2 minutes and 51 seconds.

On the ARC-AGI-3 benchmark, Astra scored 99.9%, far ahead of Anthropic's Claude Opus 5. ARC-AGI-3 is an AGI benchmark that measures the ability to navigate unfamiliar environments, work out the rules and achieve a goal. Astra also surpassed Anthropic's latest model on coding benchmarks such as Terminal-Bench 4.0 and DeepSWE v1.1. OpenAI President Greg Brockman said, "Looking back later, people will think this was the era, that this was the conversation about AGI models."

Astra drew particular attention as the first OpenAI model to receive a "Critical" security rating. That means the model has reached a level at which it can find and exploit vulnerabilities on its own, without guidelines from humans. Its ability to hide or control its reasoning process to evade human oversight has also grown stronger.

But because the model has the most advanced security capabilities, the risk of misuse is considerable. OpenAI said that on July 21 this year, GPT-5.6 Sol, which was undergoing internal testing, and some undisclosed AI models slipped out of control and carried out hacking. Apparently mindful of such concerns, OpenAI simultaneously unveiled a cybersecurity defense support program called Daybreak for Frontline Defenders.

The high cost is another weak point. Astra is priced at $10 per million input tokens and $50 per million output tokens, 2.5 times more expensive than GPT-5.6 Sol, the previous top-end model. On some coding benchmarks, it also lagged behind Claude Opus 5 and Claude Fable 5.

Original reporting by Kim Chang-young for Seoul Economic Daily.

AI-translated from Korean. Quotes from foreign sources are based on Korean-language reports and may not reflect exact original wording.

Watch · Seoul Economic Daily

More →
2:39

AI KEY

Preview
Korean Corporate Intelligence HubKOSPI · KOSDAQ · 12 sectors

A live, cap-weighted view of every KOSPI and KOSDAQ sector, with same-day Korean reporting distilled by company — built for foreign investors, correspondents and analysts who need to scan Korea before the next session.

Korea Company Atlas

Preview
Market Ontology · The Feedback LoopKFTC 2025 · 92 groups · 121,954 articles

An English ontology of the Korean market — how companies, the media, the government and the National Assembly move each other in a loop. Korea's named controlling persons and designated business groups are a mechanism, not a risk to be priced blind.

SIGNAL

Now live
English Edition · Capital MarketsM&A · IPO · PE · Fund Flows

SIGNAL English Edition is live — Korea's deal desk reporting in English. M&A, IPOs, private equity and fund flows, covered daily for global institutional investors. Browse free; subscriber-only scoops at the 50% intro rate.