Attention: You are using an outdated browser, device or you do not have the latest version of JavaScript downloaded and so this website may not work as expected. Please download the latest software or switch device to avoid further issues.

News > Attacks & Threats > OpenAI Adds Controls That Should've Been There Already

OpenAI Adds Controls That Should've Been There Already

The new AI security controls follow the Hugging Face incident last month, though experts say many of these additions should have been in place prior to the frontier models escaping.

OpenAI has committed to a number of security and guardrail improvements in the wake of an incident last month where cutting edge models inadvertently breached AI application store Hugging Face during a cyber capability benchmark exercise. Yet many experts note that of the newly announced controls appear less like groundbreaking safeguards and more like measures that should already have been in place for testing models with advanced cyber capabilities.

In response to this incident in which a model went rogue, OpenAI implemented sweeping changes. But it's not just the Hugging Face incident; OpenAI noted in an Aug. 18 blog post that preliminary evidence suggests its upcoming Astra model "may meet the Critical cybersecurity capability threshold under our Preparedness Framework." More here

Similar Stories

A purported ad-blocker exfiltrates reams of sensitive information and benefits from having Google's stamp of approval despite researcher warnings. More...

The "third-party[.]com" domain, commonly used as a documentation placeholder, has been observed serving a ClickFix lure … More...

Microsoft seized 50 websites and disabled more than 150 domains as part of a coordinated disruption effort against a phi… More...

A process parameter-poisoning technique evades EDR by injecting code into process initialization structures without usin… More...

MFA is essential, but it cannot replace OAuth governance, least-privilege scopes, consent monitoring, and rapid revocati… More...

Have your say

 

News Categories

A purported ad-blocker exfiltrates reams of sensitive information and benefits from having Google's stamp of approval despite researcher warnings. More...

The "third-party[.]com" domain, commonly used as a documentation placeholder, has been observed serving a ClickFix lure … More...

Microsoft seized 50 websites and disabled more than 150 domains as part of a coordinated disruption effort against a phi… More...

A process parameter-poisoning technique evades EDR by injecting code into process initialization structures without usin… More...

MFA is essential, but it cannot replace OAuth governance, least-privilege scopes, consent monitoring, and rapid revocati… More...

image

Contact Us

Security Interest Group Switzerland
c/o Bridge Head AG
Sulzbergstrasse 34
5430 Wettingen
Switzerland

Follow Us

This website is powered by
ToucanTech