The AI safety test is becoming a safety risk

Over the past few months, AI agents undergoing cybersecurity evaluations have escaped their boundaries, accessed the internet, and, in some cases, hacked into real-world systems. The incidents have involved models from OpenAI, Anthropic, Meta, and most recently, Chinese AI lab Moonshot AI, with testing conducted by several different organizations including a cyber evaluation startup called…

Read More
Security researchers scanned the Polish web and found courts, hospitals, and airports at risk of hacks

Security researchers scanned the Polish web and found courts, hospitals, and airports at risk of hacks

Two Polish security researchers wanted to find out how vulnerable their country’s internet was to potential cyberattacks, and quickly found that thousands of public agencies and websites were at risk of being hacked. At the Def Con cybersecurity conference in Las Vegas on Friday, security researchers Robert Kruczek and Kamil Szczurowski said they wanted to…

Read More
PSA: Your Claude shared chats and Artifacts may have ended up on Google

PSA: Your Claude shared chats and Artifacts may have ended up on Google

An untold number of Claude chats and Artifacts — the interactive mini apps and documents users can build inside Claude — were found publicly searchable on Google over the weekend, after Reddit users discovered that typing search operators like “site:claude.ai/share” into Google surfaced a long list of shared conversations. Some reportedly contained health records, private…

Read More