Ptechhub
  • News
  • Industries
    • Enterprise IT
    • AI & ML
    • Cybersecurity
    • Finance
    • Telco
  • Brand Hub
    • Lifesight
  • Blogs
No Result
View All Result
  • News
  • Industries
    • Enterprise IT
    • AI & ML
    • Cybersecurity
    • Finance
    • Telco
  • Brand Hub
    • Lifesight
  • Blogs
No Result
View All Result
PtechHub
No Result
View All Result

GPT-6 Astra Scores 100% on ExploitBench as OpenAI Blocks PoC Exploit Requests

The Hacker News by The Hacker News
September 4, 2026
Home Cybersecurity
Share on FacebookShare on Twitter


OpenAI on Thursday officially unveiled GPT‑6 Astra, which it described as the “world’s most intelligent and aligned model.”

The development comes days after the artificial intelligence (AI) company said the model had reached the “Critical” cybersecurity capability threshold under its Preparedness Framework.

“Astra is state-of-the-art on computer use, browsing, software engineering, cybersecurity, science, and professional work. Astra saturates FrontierMath Tier 4 with a 98% score,” OpenAI said. “Astra also saturates ARC-AGI-3 with a 99.9% score and ExploitBench with a 100% score. It also sets a new frontier on computer and browser use, handling the most demanding professional work with unmatched speed, accuracy, and judgment.”

The model is currently rolling out to a small set of organizations and is expected to be available to all ChatGPT Plus, Pro, Business, and Enterprise users, as well as through the OpenAI API, Microsoft Azure, and Amazon Web Services (AWS) Bedrock.

On ExploitBench, which evaluates a model’s ability to turn known software vulnerabilities into working exploits, Astra achieved a perfect score of 100%, as opposed to 78.5% for GPT‑5.6 Sol, its previous frontier cyber-capable model.

OpenAI said the model also achieves substantially higher arbitrary code-execution rates than GPT‑5.6 Sol when testing its exploit development capabilities using flaws from the previous three months between July and August 2026. This included two zero-day vulnerabilities in unspecified software.

Astra is also equipped to use previously unknown vulnerabilities to achieve code execution in hardened browsers and develop privilege-escalation exploits for hardened operating-systems, if allowed to run without any safeguards.

Given the dual-use nature of these tools – the capabilities that can help defenders find weaknesses faster can also be abused by bad actors to exploit them more easily – OpenAI said the version of Astra being released is limited to secure code review and patching, while refusing to comply with prompts related to creating proof-of-concept (PoC) exploits for vulnerabilities.

“Through OpenAI Daybreak⁠, we plan to expand access and roll out less restrictive safeguards in the coming weeks,” the AI upstart said. “This will enable more defensive workflows, including vulnerability and proof-of-concept validation, malware analysis, and detection engineering.”

To address concerns about model misuse, OpenAI said it has included stronger model robustness to better tackle jailbreaks, more context to its monitoring systems, and extra safeguards to help detect and contain misalignment.

Astra is also “more likely” to operate within the confines set by the user and implied by its environment, although it warned the safety checks can sometimes interrupt legitimate work, including defensive cybersecurity, at which point, the user will be prompted to review the action before continuing.

“In sensitive environments, Astra proceeds with care commensurate with its risk,” the company added. “In an evaluation of computer use tasks adversarially selected to elicit misbehavior, Astra was more successful at avoiding unintended consequences. Running with additional security measures offered by default yielded even stronger performance.”

The release of Astra comes as OpenAI launched a new initiative designed to provide subsidized access to its models, hands-on training, and technical assistance to critical infrastructure sectors, including water systems, electricity providers, state and local governments, banks, non-profits, open-source maintainers, and organizations with limited security resources.

The global project, called Daybreak for Frontline Defenders, aims to commit $1 billion to help defenders use frontier AI cyber capabilities to safeguard essential services against cyber attacks. In tandem, the company has announced a new pilot with the U.S. Multi-State Information Sharing and Analysis Center (MS-ISAC) to equip an initial group of public sector and water system defenders with Daybreak access, guided training, and hands-on assistance.

“We have a defender’s window: a narrowing opportunity to use AI to close security gaps before attackers seize them,” OpenAI said. “Our role is to help put powerful tools in defenders’ hands so they can protect the systems, and the people, they are responsible for.”



Source link

The Hacker News

The Hacker News

Next Post

Google Releases Chrome Update to Patch Actively Exploited V8 Zero-Day

Recommended.

Scalefusion Introduces Programmable Custom Properties

Scalefusion Introduces Programmable Custom Properties

March 11, 2026
Huawei Cloud onthult Wereldwijd beleid voor verkooppartners voor 2026 gericht op gedeeld succes in het tijdperk van AI

Huawei Cloud onthult Wereldwijd beleid voor verkooppartners voor 2026 gericht op gedeeld succes in het tijdperk van AI

January 24, 2026

Trending.

AWS, Google, Oracle, Microsoft Top Gartner’s Cloud AI Infrastructure List For 2026

AWS, Google, Oracle, Microsoft Top Gartner’s Cloud AI Infrastructure List For 2026

July 29, 2026
Cloud Market Share Q1 2026: AWS, Microsoft, Google Battling In AI Era

Cloud Market Share Q1 2026: AWS, Microsoft, Google Battling In AI Era

May 4, 2026

Goldman Sachs picks China stocks poised to benefit from a new wave of AI-related hardware exports

August 16, 2026
Anthropic lost control of Claude in latest AI cyber blunder | Computer Weekly

Anthropic lost control of Claude in latest AI cyber blunder | Computer Weekly

July 31, 2026
Sohu.com to Report Second Quarter 2026 Financial Results on August 10, 2026

Sohu.com to Report Second Quarter 2026 Financial Results on August 10, 2026

July 31, 2026

PTechHub

A tech news platform delivering fresh perspectives, critical insights, and in-depth reporting — beyond the buzz. We cover innovation, policy, and digital culture with clarity, independence, and a sharp editorial edge.

Follow Us

Industries

  • AI & ML
  • Cybersecurity
  • Enterprise IT
  • Finance
  • Telco

Navigation

  • About
  • Advertise
  • Privacy & Policy
  • Contact

Subscribe to Our Newsletter

  • About
  • Advertise
  • Privacy & Policy
  • Contact

Copyright © 2025 | Powered By Porpholio

No Result
View All Result
  • News
  • Industries
    • Enterprise IT
    • AI & ML
    • Cybersecurity
    • Finance
    • Telco
  • Brand Hub
    • Lifesight
  • Blogs

Copyright © 2025 | Powered By Porpholio