-
ExploitBench: Claude Mythos Preview and GPT-5.5 Develop Real Browser Exploits Autonomously
Anthropic
research
-
Anthropic Expands Project Glasswing to ~200 Partners, Grants Mythos Preview Access for Critical Infrastructure
Anthropic
industry
-
Trump Signs AI Executive Order Requiring 30-Day Voluntary Pre-Release Government Review
industry
-
US Lifts Export Controls on Anthropic's Claude Fable 5 and Mythos 5
Anthropic
industry
-
OpenAI unveils GPT-Red, an internal automated red-teaming model for prompt-injection defense
OpenAI
research
-
OpenAI says pre-release models broke out of test sandbox and breached Hugging Face
OpenAI
research
-
OpenAI Launches Daybreak AI Cybersecurity Initiative with GPT-5.5 Models
OpenAI
tools
-
OpenAI Rolls Out Lockdown Mode to Block Prompt-Injection Exfiltration in ChatGPT
OpenAI
tools
-
Claude Code v2.1.177: Fable 5 Forced Fallback to Opus 4.8, Bedrock Cache Fix, Security Patch
Anthropic
tools
-
NVIDIA SkillSpector: Open-Source Security Scanner for AI Agent Skills
NVIDIA
tools
-
xAI Open-Sources Grok Build Coding Agent Under Apache 2.0 After Privacy Incident
xAI
tools
-
GitHub MCP Server: Secret Scanning GA and Dependency Scanning Public Preview
GitHub
tools
-
Claude Code 2.1.178: Parameterized Permission Rules and Nested Skills
Anthropic
tools
-
Anthropic Proposes Industry-Wide Cyber Jailbreak Severity Scale
Anthropic
research
-
Claude Code v2.1.207: Auto Mode GA on Bedrock/Vertex/Foundry, Security Fixes
Anthropic
tools
-
BadHost (CVE-2026-48710): Host-Header Auth Bypass in Starlette Exposes vLLM, LiteLLM, and MCP Servers
tools
-
Anthropic's Mythos Model Found Vulnerabilities in Classified US Government Systems Within Hours
Anthropic
industry
-
GLM-5.2: Zhipu AI's MIT-Licensed 744B MoE Coding Model Raises Cybersecurity Concerns
Zhipu AI / Z.ai
models-llm
-
Anthropic launches Claude Security in public beta for enterprise customers
Anthropic
tools
-
Fake OpenAI Repo Hits #1 Trending on Hugging Face with 244K Downloads, Delivers Infostealer
tools
-
Anthropic Accuses Alibaba of Largest Known Claude Distillation Attack: 28.8M Conversations
Anthropic
industry
-
OpenClaw v2026.5.12-beta.4/5/6: Security Hardening and Multi-Platform Messaging Fixes
tools
-
Claude Code v2.1.160: Security Prompts Before Writing Shell Startup Files and Build-Tool Configs
Anthropic
tools
-
Claude Code v2.1.162: Security Fix for OAuth Credential Leak, Parallel Tool Call Isolation
Anthropic
tools
-
Claude Code v2.1.166: Fallback Model Config, Expanded Deny-Rule Globs, Cross-Session Security
Anthropic
tools
-
Claude Code v2.1.187: Sandbox Credential Isolation and Remote MCP Hang Fix
Anthropic
tools
-
Claude Code v2.1.193: Shell Classifier Expansion, OTel Response Logging, Live Path Autocomplete
Anthropic
tools
-
OpenAI Codex CLI v0.142.2: Default MCP Tool Search, macOS Proxy Support, PowerShell Safety
OpenAI
tools
-
Claude Code v2.1.196: Org Default Models and MCP Security Fix for Untrusted Repos
Anthropic
tools
-
OpenAI Codex v0.142.5: Security Fix for Trace Log WebSocket Exposure
OpenAI
tools
-
GitHub Copilot CLI Drops PAT Requirement in Actions; Agent Session Streaming in Preview
GitHub
tools
-
Claude Code v2.1.205: Transcript-Tampering Guard and Session Reliability Fixes
Anthropic
tools
-
Claude Code v2.1.207: Auto Mode GA on Cloud Providers, Security Fix
Anthropic
tools
-
GitHub Copilot Adds Prompt-Injection Detection in CodeQL 2.26.0
GitHub
tools
-
Claude Code v2.1.210–211: Bidirectional-Override Security Patch and Subagent Stream Flag
Anthropic
tools
-
Claude Code v2.1.214–v2.1.215: permission-bypass fixes, EndConversation tool, /verify no longer auto-runs
Anthropic
tools
-
Safeguards Based on Copyable Context Cannot Provide Reliable Safety for LLMs
research
-
OpenAI Ships codex-zsh v0.1.0: Versioned Patched zsh Binary for Codex Sandbox
OpenAI
tools
-
OpenAI Codex CLI 0.142.5: Trace-Log Privacy Fix
OpenAI
tools