Attention: You are using an outdated browser, device or you do not have the latest version of JavaScript downloaded and so this website may not work as expected. Please download the latest software or switch device to avoid further issues.
| 31 Jul 2026 | |
| Written by Gabi Gerber | |
| Attacks & Threats |
Major frontier AI vendors — including Anthropic, Google, and OpenAI — need to rein in the harnesses they wrap around their large language modules, to limit security weaknesses created by software components that are too trusting of each other.
That's the word from researchers at AI penetration testing firm Novee Security, who were able to use Google's AI agent to execute a supply chain attack and write to its own repository on GitHub, says Elad Meged, a founding team and security researcher at the company. The team also found issues in Anthropic's and OpenAI's AI agents by exploiting misalignments in the trust between elements to enable attacks. More here
Dark Reading Confidential Episode 20: Expert Rich Mogull reflects on lessons cyber teams should pull from the OpenAI agent's attack on Hugging Face. More...
New research shows how attackers can use security alerts and blocked events to manipulate and hijack AI agents. More...
A study of more than 6,000 patches found that even working patches can introduce new bugs, break something else, or are … More...
Advanced LLMs escaped their sandboxes while attempting to achieve a non-malicious benchmark test objective. More...
Ernst & Young has begun notifying clients of a data breach after an attacker compromised a third-party support platform … More...
Dark Reading Confidential Episode 20: Expert Rich Mogull reflects on lessons cyber teams should pull from the OpenAI agent's attack on Hugging Face. More...
New research shows how attackers can use security alerts and blocked events to manipulate and hijack AI agents. More...
A study of more than 6,000 patches found that even working patches can introduce new bugs, break something else, or are … More...
Advanced LLMs escaped their sandboxes while attempting to achieve a non-malicious benchmark test objective. More...
A myriad of software makes up the typical AI harness, and trust issues between the components can create concerning atta… More...