Maybe feeling left out from the questionable hype train of “Our AI models can’t be trusted”, Anthropic has released reports that their Claude model has “reached the ...
Meta on Thursday said that one of its artificial intelligence models accessed the internet on its own and hacked another ...
On March 24, 2026, developers building AI applications with LiteLLM — a Python package with 95 million monthly downloads — ...
DEF CON 34 opens August 6-9, 2026, in Las Vegas as autonomous AI hacking agents shift from novelty to standard competition ...
Zenity has disclosed the details of two AI browser hacking techniques targeting Claude in Chrome and ChatGPT Atlas.
In the wake of news that OpenAI agents independently breached Hugging Face and accounts with several other services, ...
The AI models tested carried out “unsanctioned” actions — including hacking a website and attempting to inject harmful code into software — reinforcing fears that neither the creators nor seasoned ...
The UK AI Security Institute says OpenAI’s and and Anthropic’s models engaged in deceptive behavior and harmful activity ...
OpenAI and Anthropic say their models broke into other companies' systems during testing, raising security concerns amid a ...
Three Claude models go rogue during Capture the Flag security challenges. Here's the trail of damage each left behind.
Barely a week after OpenAI admitted its models attacked Hugging Face, Anthropic is owning up to Claude’s own real-life hacking attempts.
NPR's HOST HOST speaks with Thomas Wolf, co-founder and chief science officer for Hugging Face, an AI company, discusses the recent hacking incident involving ChatGPT.