OpenAI says its upcoming Astra model crosses a 'Critical' cyber threshold

Started by QuantumToken65, Today at 06:16 AM

Previous topic - Next topic

0 Members and 1 Guest are viewing this topic.

Topic: OpenAI says its upcoming Astra model crosses a 'Critical' cyber threshold   Views(Read 51 times)
Active members in this topic:
QuantumToken65(1)

QuantumToken65

OpenAI said on September 1st that its upcoming AI model Astra is the first of its models to exceed the company's Critical cybersecurity capability threshold under its internal Preparedness Framework. The company said Astra can identify previously unknown security vulnerabilities and develop exploitation techniques without step by step human guidance, placing it in the most advanced risk category OpenAI tracks for potential severe harm

OpenAI said it still plans to release Astra soon, but that access to its most advanced cybersecurity capabilities will initially be limited to a small group of alpha testers, described as individuals and organizations responsible for protecting critical digital infrastructure, including elements of the US government and companies in OpenAI's trusted access program. A broader pool of users will later get access for defensive cybersecurity purposes through OpenAI's Daybreak Blue program once the company is confident the model's access controls are properly calibrated

The company said Astra's release had already been delayed several weeks after everything was paused in the wake of the earlier Hugging Face breach incident, and that although Astra itself wasn't involved in that specific incident, lessons from it directly informed the additional guardrails now being applied. Curious what people think about this kind of staged, restricted rollout as AI models increasingly cross thresholds that make them individually capable of both defending against and enabling serious cyberattacks


Save money on everyday spending Free cashback on thousands of retailers
View offer