On March 24, 2026, developers building AI applications with LiteLLM — a Python package with 95 million monthly downloads — ...
OpenAI and Anthropic disclosed incidents in which frontier AI models carried out unauthorized hacking-related actions. OpenAI said its models exploited a zero-day vulnerability and breached part of ...
Unravelling the hype behind IT for creating useful CIO strategies. Anyone considering what guide rails need to be in place to protect us from rogue AI behaviour, need to read Anthropic’s postmortem of ...
Anthropic says Claude models breached three real companies during cyber tests, exposing serious gaps in AI evaluation ...
Dependency confusion is a supply chain issue that affects how package managers choose where to download a dependency from. If your build or developer tooling can see both a private package registry ...
OpenAI and Anthropic's July AI agent breaches revive Nick Bostrom's paperclip maximizer thought experiment and instrumental convergence theory.
AI safety federal investigation call from 15 organizations reaches President Trump on July 30, as Anthropic disclosed that ...
Anthropic said the OpenAI event spurred its engineers to review similar cybersecurity evaluations by Claude models. The audit ...
A figurine in front of the logo of the AI assistant "Claude" built by the US artificial intelligence safety and research company Anthropic during a photo session in Paris on February 13, 2026. Joel ...
Anthropic says Claude models breached three organizations after escaping a misconfigured cyber evaluation environment run with Irregular.
Wrote and published malware during tests, which is apparently OK because leaky test environments were the real problem ...
The AI model repeatedly tried to obtain funds for a phone number to create an account before eventually publishing a ...
Some results have been hidden because they may be inaccessible to you
Show inaccessible results