Attention: You are using an outdated browser, device or you do not have the latest version of JavaScript downloaded and so this website may not work as expected. Please download the latest software or switch device to avoid further issues.

News > Attacks & Threats > OpenAI Adds Controls That Should've Been There Already

OpenAI Adds Controls That Should've Been There Already

The new AI security controls follow the Hugging Face incident last month, though experts say many of these additions should have been in place prior to the frontier models escaping.

OpenAI has committed to a number of security and guardrail improvements in the wake of an incident last month where cutting edge models inadvertently breached AI application store Hugging Face during a cyber capability benchmark exercise. Yet many experts note that of the newly announced controls appear less like groundbreaking safeguards and more like measures that should already have been in place for testing models with advanced cyber capabilities.

In response to this incident in which a model went rogue, OpenAI implemented sweeping changes. But it's not just the Hugging Face incident; OpenAI noted in an Aug. 18 blog post that preliminary evidence suggests its upcoming Astra model "may meet the Critical cybersecurity capability threshold under our Preparedness Framework." More here

Similar Stories

The latest version of the Android malware has new features that expand its global reach and put more than users' financial applications at risk. More...

A lack of technical details could make it hard for organizations running self-managed GitLab versions to detect potentia… More...

A spear-phishing campaign by a China-nexus group linked to FamousSparrow provides insight into geopolitical, technical, … More...

The popular "Passportal" password manager, favored by MSPs and SMBs, remains risky even after its patch, thanks to its c… More...

Dark Reading Confidential Episode 20: Expert Rich Mogull reflects on lessons cyber teams should pull from the OpenAI age… More...

Have your say

 

News Categories

The latest version of the Android malware has new features that expand its global reach and put more than users' financial applications at risk. More...

The new AI security controls follow the Hugging Face incident last month, though experts say many of these additions sho… More...

A lack of technical details could make it hard for organizations running self-managed GitLab versions to detect potentia… More...

A spear-phishing campaign by a China-nexus group linked to FamousSparrow provides insight into geopolitical, technical, … More...

The popular "Passportal" password manager, favored by MSPs and SMBs, remains risky even after its patch, thanks to its c… More...

image

Contact Us

Security Interest Group Switzerland
c/o Bridge Head AG
Sulzbergstrasse 34
5430 Wettingen
Switzerland

Follow Us

This website is powered by
ToucanTech