Tag: Security
-
AI · · September 11, 2026
-
AI · · September 7, 2026
-
AI · · September 3, 2026
-
Anthropic details how its models broke out of test sandboxes (anthropic.com)AI · · September 1, 2026
-
Simon Willison maps out what ChatGPT Work can actually do (simonwillison.net)AI · · September 1, 2026
-
The Hugging Face hack postmortem: agents that coordinated to cheat (thezvi.wordpress.com)AI · · August 31, 2026
-
A rumour of a bug is now enough for AI to build the exploit (anil.recoil.org)AI · · August 29, 2026
-
A prompt injection attack turns Claude Code's auto mode against itself (simonwillison.net)AI · · August 28, 2026
-
smolvm runs untrusted code in fast, hardware-isolated micro-VMs (simonwillison.net)AI · · August 22, 2026
-
AI · · August 18, 2026
-
Inside the gray market for reselling AI API credits (vectoral.com)AI · · August 17, 2026
-
Researchers pulled hidden reasoning out of Claude, GPT, and Gemini APIs (simonwillison.net)AI · · August 13, 2026
-
Claude Code turns on auto mode by default (claude.com)AI · · August 10, 2026
-
The real lesson from AI models hacking during tests: nobody is ready (interconnects.ai)AI · · August 10, 2026
-
OpenAI slows Astra after cyber tests approach a critical threshold (techcrunch.com)Security · · August 9, 2026
-
Now three labs have had models attack real systems during tests (simonwillison.net)AI · · August 7, 2026
-
Atlassian Rovo can be tricked into leaking Jira and Confluence data (promptarmor.com)AI · · August 6, 2026
-
AI · · August 5, 2026
-
JFrog finds 54 of 55 SQLite CVEs were fabricated (research.jfrog.com)AI · · August 4, 2026
-
AI · · July 30, 2026
-
Claude found new weaknesses in two cryptographic schemes (anthropic.com)AI · · July 28, 2026
-
A forensic timeline of how an AI agent breached Hugging Face (huggingface.co)Security · · July 28, 2026
-
Security · · July 27, 2026
-
Google's small Gemini Flash Cyber outfinds larger models on bugs (deepmind.google)AI · · July 23, 2026
-
AI models escaped a test sandbox and hacked Hugging Face (huggingface.co)Security · · July 22, 2026
-
xAI open-sources its coding agent after a privacy backlash (simonwillison.net)AI · · July 19, 2026
-
An AI auditor found a critical bug in OpenVM, but only with the right crypto context (blog.zksecurity.xyz)AI · · July 18, 2026
-
Capital One open-sources VulnHunter, an agentic bug finder that checks its own work (capitalone.com)AI · · July 18, 2026
-
A honeypot site tricked Claude's web_fetch into leaking user data (simonwillison.net)Security · · July 16, 2026
-
AI · · July 16, 2026
-
Microsoft patches a record 570 bugs and blames AI for finding them (techcrunch.com)AI · · July 16, 2026
-
Anthropic drafts a severity scale for AI jailbreaks (anthropic.com)AI · · July 4, 2026
-
Security · · June 29, 2026
-
AI · · June 27, 2026
-
AI · · June 23, 2026
-
AI · · June 23, 2026
-
DeepMind plans to treat its own AI agents as insider threats (deepmind.google)AI · · June 21, 2026
-
DeepMind treats a misbehaving AI agent like an insider threat (deepmind.google)AI · · June 20, 2026
-
Research agents leak secrets through their own search queries (huggingface.co)AI · · June 19, 2026
-
AI · · June 16, 2026
-
AI · · June 11, 2026
-
Simon Willison's MicroPython sandbox is 362 KB of WASM with fuel-based CPU limits (simonwillison.net)AI · · June 9, 2026
-
OpenAI ships a Lockdown Mode against prompt injection exfiltration (simonwillison.net)Security · · June 7, 2026
-
Security · · June 3, 2026
-
AI · · June 2, 2026
-
Meta AI handed out Instagram accounts when politely asked (simonwillison.net)AI · · June 2, 2026
-
How Anthropic contains Claude across its products (simonwillison.net)Engineering · · May 31, 2026
-
AI-assisted security reports to curl are arriving at 4-5x the 2024 rate (simonwillison.net)Security · · May 30, 2026
-
Security · · May 30, 2026
-
AI · · May 27, 2026
-
Microsoft Copilot Cowork can be tricked into leaking files (simonwillison.net)AI · · May 26, 2026
-
AI · · May 24, 2026
-
AI · · May 23, 2026
-
Gemini Spark wires an agent into your inbox, and Simon Willison is worried (simonwillison.net)AI · · May 22, 2026
-
AI · · May 22, 2026
-
Frontier AI broke competitive CTF (kabir.au)Security · · May 17, 2026
-
Security · · May 12, 2026
-
Mozilla says it fixed 423 Firefox security bugs in one month with AI help (simonwillison.net)Security · · May 7, 2026
-
Jack Clark sorts attacks on AI agents into six kinds (importai.substack.com)AI · · April 13, 2026
-
Safetensors moves to the PyTorch Foundation, with neutral governance (huggingface.co)Engineering · · April 8, 2026
-
Security · · April 7, 2026