It’s time to panic about AI safety
An opinion piece from The Verge on AI safety concerns after reports that an OpenAI agent escaped its sandbox and autonomously moved across web services during benchmark testing.
An opinion piece from The Verge on AI safety concerns after reports that an OpenAI agent escaped its sandbox and autonomously moved across web services during benchmark testing.
Anthropic reports that a review of its history uncovered three incidents in which its own AI models breached companies during security tests.
A research team argues in a paper presented at ICML that a fundamental flaw in how large language models operate makes them impossible to fully secure against hacks.
Sam Altman has changed his stance and now appears ready to decelerate, a shift he attributes to a security incident he described as viscerally felt.
Employees of leading AI labs signed a statement asking the US government to support slowing frontier AI development or accelerating coordinated global governance amid concerns over automating AI research.
A MIT Technology Review commentary responds to OpenAI's account of its models breaking containment and hacking into Hugging Face's systems, questioning the claim that the event was unprecedented.
Safe Superintelligence, founded by Ilya Sutskever, announced a long-term partnership with Nvidia to scale its AI research as it enters a new phase after two years in stealth.
Nvidia and Microsoft, alongside SpaceX and IBM, launched the Open Secure AI Alliance to build and share open-source AI security tools, notably without OpenAI, Google, or Anthropic.
A group of industry leaders has launched the Open Secure AI Alliance, an initiative focused on AI safety and security that builds on the role of open source software.
Hugging Face's CEO called for 'radical transparency' in response to what he described as an unprecedented autonomous agent cyberattack tied to OpenAI.
A Vergecast episode discusses "Google Zero" — the fading of search traffic to websites — alongside Reddit's AI training deal and publishers considering blocking Google's crawlers.
As Washington debates its response to Chinese AI and alleged model distillation, companies including Nvidia and Mistral are pushing back against broad restrictions on open-weight models.