
OpenAI said on Sept. 1 that it is preparing to release its latest advanced artificial intelligence model, Astra, after introducing stronger safeguards following a cybersecurity incident involving other models under development.
The San Francisco-based company paused some model development for two weeks this summer after two AI models being tested were involved in a security breach affecting software platform Hugging Face. OpenAI said Astra was not involved in that incident but has since strengthened its security measures ahead of its release.
According to OpenAI, Astra has been trained to more reliably reject harmful cybersecurity requests and comply with safety restrictions. The company has also added protections against misuse and monitoring systems designed to detect and stop potentially unauthorised activity.
OpenAI has classified Astra as meeting a “critical cybersecurity threshold”, meaning it believes the model could identify and exploit vulnerabilities in computer systems. The company said Astra is the first model it has designated at this level, requiring enhanced safeguards throughout development and before its public release.
Access to some of Astra’s capabilities will initially be restricted, while its most advanced functions will be provided to a limited group of early testers, OpenAI said.
Concerns over the cybersecurity risks posed by increasingly capable AI models have grown following incidents involving systems developed by OpenAI and rival Anthropic. Anthropic recently said its models gained unauthorised access to systems belonging to three organisations during testing intended to prevent contact with real-world systems.
More than 100 organisations, including OpenAI and Anthropic, recently signed an open letter calling for stronger global cyber defences against AI-powered threats. The groups warned that AI-enabled cyberattacks could become more widespread and sophisticated as models continue to improve.
In June, US President Donald Trump signed an executive order establishing a voluntary review process under which the government would receive early access to new AI models to assess potential security risks. Although a final framework was expected by Aug. 1, the White House has not publicly released it.
OpenAI said it is nevertheless following the voluntary framework as it prepares to launch Astra. – TS/ERMD
READ MORE
From Slide Rules to Artificial Intelligence — An Engineer’s 60-Year Journey
Engr. Farooq Mehboob reflects on six decades of engineering, from manual calculations and slide rules…
Engineering, Public Service and Nation Building: The Journey of Engr. Mukhtiar A. Shaikh
From the classrooms of Dawood Engineering College to leadership roles in public service, industry, and…
Building a Stronger Nation: Tahir Sultan on Engineering, Governance, and Economic Revival
Engineering Review: In the aftermath of the Middle East war, Pakistan’s fault lines and weaknesses…
Anthropic Researcher Resigns Over AI Safety Concerns
An Anthropic researcher has resigned from the artificial intelligence (AI) company, warning that leading AI…
Qaiser Ahmed Sheikh Calls for Turning Economic Knowledge into Practical Solutions
Federal Minister for Board of Investment Qaiser Ahmed Sheikh has called for greater efforts to…
Apple Unveils $1,999 Folding iPhone Duo in Major Flagship Redesign
Apple has unveiled the iPhone Duo, a $1,999 foldable smartphone that marks the company’s biggest…
