Attention: You are using an outdated browser, device or you do not have the latest version of JavaScript downloaded and so this website may not work as expected. Please download the latest software or switch device to avoid further issues.

News > Attacks & Threats > OpenAI Adds Controls That Should've Been There Already

OpenAI Adds Controls That Should've Been There Already

The new AI security controls follow the Hugging Face incident last month, though experts say many of these additions should have been in place prior to the frontier models escaping.

OpenAI has committed to a number of security and guardrail improvements in the wake of an incident last month where cutting edge models inadvertently breached AI application store Hugging Face during a cyber capability benchmark exercise. Yet many experts note that of the newly announced controls appear less like groundbreaking safeguards and more like measures that should already have been in place for testing models with advanced cyber capabilities.

In response to this incident in which a model went rogue, OpenAI implemented sweeping changes. But it's not just the Hugging Face incident; OpenAI noted in an Aug. 18 blog post that preliminary evidence suggests its upcoming Astra model "may meet the Critical cybersecurity capability threshold under our Preparedness Framework." More here

Similar Stories

Threat actors behind the "Phantom Deal" campaign are studying companies in extreme detail, aiming to dupe midlevel employees into initiating large financial transfers. More...

The "Spring Ring" operation aims to compromise users of the collaboration suite to remotely access their sessions, sprea… More...

Microsoft has disclosed details of a new ClickFix variant, dubbed TerminalFix, that aims to trick users into running a m… More...

OpenAI's Hugging Face attack postmortem shows agents don't care about rules — they need strong controls. More...

An untold number of ZBT routers sold around the world as white-label products come with several implants built by the ma… More...

Have your say

 

News Categories

Threat actors behind the "Phantom Deal" campaign are studying companies in extreme detail, aiming to dupe midlevel employees into initiating large financial transfers. More...

The "Spring Ring" operation aims to compromise users of the collaboration suite to remotely access their sessions, sprea… More...

Microsoft has disclosed details of a new ClickFix variant, dubbed TerminalFix, that aims to trick users into running a m… More...

OpenAI's Hugging Face attack postmortem shows agents don't care about rules — they need strong controls. More...

An untold number of ZBT routers sold around the world as white-label products come with several implants built by the ma… More...

image

Contact Us

Security Interest Group Switzerland
c/o Bridge Head AG
Sulzbergstrasse 34
5430 Wettingen
Switzerland

Follow Us

This website is powered by
ToucanTech