Breaking
CNNepali woman rescues two schoolchildren from landslideDEZelensky warns against drone attacks on Russian airspaceITElderly man found dead at home in Imola, wife in a state of confusionSEThe US attacks Iranian targets after suspected mine laying in the Strait of HormuzINTLWhistleblower Warns Trump Postal Service Plan Could Disenfranchise Millions of VotersFRSerious accidents in the Pyrénées-Orientales and Bouches-du-Rhône: two deaths and a police officer in a comaDEInflation in the euro area rises to 3.3 percent – ECB has to raise interest ratesCNPedrey Ng Wing-lam overcomes perfectionism to pursue Olympic table tennis dreamDEVenezuelan parliament approves oil deal with USRUManturov says presidential decree on temporary management generalizes prior workCNNepali woman rescues two schoolchildren from landslideDEZelensky warns against drone attacks on Russian airspaceITElderly man found dead at home in Imola, wife in a state of confusionSEThe US attacks Iranian targets after suspected mine laying in the Strait of HormuzINTLWhistleblower Warns Trump Postal Service Plan Could Disenfranchise Millions of VotersFRSerious accidents in the Pyrénées-Orientales and Bouches-du-Rhône: two deaths and a police officer in a comaDEInflation in the euro area rises to 3.3 percent – ECB has to raise interest ratesCNPedrey Ng Wing-lam overcomes perfectionism to pursue Olympic table tennis dreamDEVenezuelan parliament approves oil deal with USRUManturov says presidential decree on temporary management generalizes prior work
BackOpenAI's Astra AI Model Crosses Critical Cybersecurity Threshold
OpenAI's Astra AI Model Crosses Critical Cybersecurity Threshold
Developing
CNBC56 minutes agoTech1 min read

OpenAI's Astra AI Model Crosses Critical Cybersecurity Threshold

Quick Look

  • OpenAI announced its upcoming AI model Astra is the first to cross its 'Critical' cybersecurity capability threshold, meaning it can find and exploit unknown security flaws without human guidance.
  • Despite planning a soon release, access to Astra's cyber capabilities will be limited to select organizations in its Daybreak coalition due to safety concerns following recent model breaches.

AI-generated summary

Why It Matters

OpenAI introduced its Preparedness Framework in 2023 to track advanced AI capabilities that could introduce risks of severe harm, with 'Critical' threshold models able to introduce unprecedented new pathways to severe harm.

Font size

OpenAI on Tuesday said its upcoming artificial intelligence model Astra is the first offering that crosses its "Critical" cybersecurity capability threshold.

The company said Astra can find previously unknown security flaws and exploit them without step-by-step guidance from humans, which means the model falls under the most advanced category of its so-called Preparedness Framework. OpenAI said it still plans to make Astra available "soon," but that access to its cybersecurity capabilities will be more limited.

OpenAI introduced its Preparedness Framework in 2023, and it serves as the company's method for "tracking and preparing for advanced AI capabilities that could introduce new risks of severe harm." In an update to the framework last year, the company outlined a "High" capability threshold, where models could amplify "existing pathways" to severe harm, and a "Critical" capability threshold, where models could introduce "unprecedented new pathways" to severe harm.

"We will share more details about our safety, security and alignment testing and evaluations in the model's System Card at launch," OpenAI said in a blog post on Tuesday.

OpenAI's security and safety practices have been under intense scrutiny after the company disclosed that two of its models escaped their training environment, accessed the open web and breached Hugging Face's systems last month. OpenAI characterized the attack as an "unprecedented cyber incident" and temporarily paused some of its internal training and research.

The company decided to delay parts of Astra's development even though the model was not involved in the Hugging Face incident. After strengthening and testing protections, OpenAI said Tuesday that it believes the model's safeguards "sufficiently minimize the risk of severe harm for release under our Preparedness Framework."

Astra's advanced cyber capabilities will be available to a select group of organizations that are part of its cybersecurity coalition called Daybreak, OpenAI said.

What to Watch

AI outlook — possibilities, not facts

  • OpenAI will release Astra's cybersecurity capabilities to the Daybreak coalition in the near future

    Very likely · Within weeks

  • OpenAI will publish detailed safety and security testing results for Astra in its System Card at launch

    Very likely · Within weeks

Open Questions

  • When exactly will Astra be released to the public?
  • Which specific organizations are part of the Daybreak cybersecurity coalition?
  • What specific safeguards has OpenAI strengthened for Astra's release?

Related Topics

This article was originally published by CNBC.

Related Stories

B.AI Crosses 2 Trillion Token Threshold in Free Access Campaign
Developing·56 minutes ago

B.AI Crosses 2 Trillion Token Threshold in Free Access Campaign

B.AI announced that its sitewide free-access campaign for AI models surpassed 2 trillion cumulative tokens processed, achieved over seven days starting August 17 with free access to DeepSeek V4 Flash, Tencent Hy3, and other models. The milestone demonstrates the platform's scalability and cost-efficient infrastructure, which uses dual-tier API routing and Web2/Web3 payment integration to reduce enterprise AI compute costs by up to 90%.

CryptoSlate
3 min read
Pyka and other companies advance autonomous fixed-wing aircraft for crop spraying and cargo delivery
Developing·1 hour ago

Pyka and other companies advance autonomous fixed-wing aircraft for crop spraying and cargo delivery

Pyka's pilotless crop-spraying planes are flying in California and Brazil, offering reduced chemical use through lower flight altitudes. The company joins Reliable Robotics and Merlin Labs in advancing autonomous fixed-wing aviation, with differing approaches to AI and certification, aiming eventually for passenger service amid stricter safety standards than self-driving cars.

BBC News
2 min read
More on this topicopenai