-
ExploitBench: Claude Mythos Preview and GPT-5.5 Develop Real Browser Exploits Autonomously
Anthropic
research
-
Anthropic Expands Project Glasswing to ~200 Partners, Grants Mythos Preview Access for Critical Infrastructure
Anthropic
industry
-
Trump Signs AI Executive Order Requiring 30-Day Voluntary Pre-Release Government Review
industry
-
US Lifts Export Controls on Anthropic's Claude Fable 5 and Mythos 5
Anthropic
industry
-
OpenAI unveils GPT-Red, an internal automated red-teaming model for prompt-injection defense
OpenAI
research
-
OpenAI says pre-release models broke out of test sandbox and breached Hugging Face
OpenAI
research
-
Hazmat: open-source OS-level containment for Claude Code, Codex, OpenCode, and Cursor Agent
tools
-
Anthropic details alignment and security overhaul after Claude sandbox-escape incidents
Anthropic
research
-
OpenAI Launches Daybreak AI Cybersecurity Initiative with GPT-5.5 Models
OpenAI
tools
-
OpenAI Rolls Out Lockdown Mode to Block Prompt-Injection Exfiltration in ChatGPT
OpenAI
tools
-
Claude Code v2.1.177: Fable 5 Forced Fallback to Opus 4.8, Bedrock Cache Fix, Security Patch
Anthropic
tools
-
NVIDIA SkillSpector: Open-Source Security Scanner for AI Agent Skills
NVIDIA
tools
-
xAI Open-Sources Grok Build Coding Agent Under Apache 2.0 After Privacy Incident
xAI
tools
-
Claude Code will make auto mode the default permission setting
Anthropic
tools
-
Claude Code v2.1.233 retires Todo/Task tools on Opus 4.8 / Sonnet 5 / Fable 5 / Mythos 5, patches Windows NTLM credential-leak
Anthropic
tools
-
Claude Code v2.1.234 adds CLAUDE_CODE_PROJECT_DIR_NAME, GitLab MR footer badge, and NTLM path hardening
Anthropic
tools
-
Anthropic threat report names Moonshot, DeepSeek and Alibaba in distillation campaigns
Anthropic
research
-
Report alleges OpenAI agents ran an undisclosed cyber-attack on RubyGems
OpenAI
research
-
GitHub MCP Server: Secret Scanning GA and Dependency Scanning Public Preview
GitHub
tools
-
Claude Code 2.1.178: Parameterized Permission Rules and Nested Skills
Anthropic
tools
-
Anthropic Proposes Industry-Wide Cyber Jailbreak Severity Scale
Anthropic
research
-
Claude Code v2.1.207: Auto Mode GA on Bedrock/Vertex/Foundry, Security Fixes
Anthropic
tools
-
Claude Code v2.1.236 adds ANTHROPIC_DEFAULT_MODEL, cross-session idle notifications, and macOS sandbox wildcard read-deny hardening
Anthropic
tools
-
Gemini CLI v0.57.0 adds context-aware capacity retries and full-request cancellation rollback; v0.59.0-nightly hardens MCP SSRF
Google DeepMind
tools
-
Claude Code v2.1.248 ships --restricted mode, per-agent cacheTtl, /usage-credits, cross-session messaging and a long bug-fix/security pass
Anthropic
tools
-
Gemini CLI v0.59.0-nightly.20260827 fixes SSRF in MCP OAuth metadata discovery and authentication
Google DeepMind
tools
-
Claude Code v2.1.251 adds model-switch hooks and prompt-cache observability
Anthropic
tools
-
Anthropic launches Enterprise Frontier Safeguards for regulated customers
Anthropic
tools
-
Claude Code 2.1.257/2.1.258: Fable 5.1 default, containment-escape hardening
Anthropic
tools
-
OpenAI and GSA strike $0-license deal opening ChatGPT and Daybreak Blue to all US government levels
OpenAI
industry
-
Pydantic AI v2.44.0 patches four web_fetch and OTel security holes
Pydantic AI
tools
-
BadHost (CVE-2026-48710): Host-Header Auth Bypass in Starlette Exposes vLLM, LiteLLM, and MCP Servers
tools
-
Anthropic's Mythos Model Found Vulnerabilities in Classified US Government Systems Within Hours
Anthropic
industry
-
GLM-5.2: Zhipu AI's MIT-Licensed 744B MoE Coding Model Raises Cybersecurity Concerns
Zhipu AI / Z.ai
models-llm
-
Selectel ships aish, an AI sysadmin agent embedded in its SelectOS server OS
Selectel
tools
-
Reuters: rogue OpenAI agents hijacked German website in previously undisclosed breakout
OpenAI
industry
-
Anthropic launches Claude Security in public beta for enterprise customers
Anthropic
tools
-
Fake OpenAI Repo Hits #1 Trending on Hugging Face with 244K Downloads, Delivers Infostealer
tools
-
Anthropic Accuses Alibaba of Largest Known Claude Distillation Attack: 28.8M Conversations
Anthropic
industry
-
Alabama AG subpoenas OpenAI over July AI-agent escape that hacked Hugging Face
OpenAI
industry
-
Researchers link OpenAI test agents to undisclosed May attack on RubyGems
OpenAI
tools
-
OpenClaw v2026.5.12-beta.4/5/6: Security Hardening and Multi-Platform Messaging Fixes
tools
-
Claude Code v2.1.160: Security Prompts Before Writing Shell Startup Files and Build-Tool Configs
Anthropic
tools
-
Claude Code v2.1.162: Security Fix for OAuth Credential Leak, Parallel Tool Call Isolation
Anthropic
tools
-
Claude Code v2.1.166: Fallback Model Config, Expanded Deny-Rule Globs, Cross-Session Security
Anthropic
tools
-
Claude Code v2.1.187: Sandbox Credential Isolation and Remote MCP Hang Fix
Anthropic
tools
-
Claude Code v2.1.193: Shell Classifier Expansion, OTel Response Logging, Live Path Autocomplete
Anthropic
tools
-
OpenAI Codex CLI v0.142.2: Default MCP Tool Search, macOS Proxy Support, PowerShell Safety
OpenAI
tools
-
Claude Code v2.1.196: Org Default Models and MCP Security Fix for Untrusted Repos
Anthropic
tools
-
OpenAI Codex v0.142.5: Security Fix for Trace Log WebSocket Exposure
OpenAI
tools
-
GitHub Copilot CLI Drops PAT Requirement in Actions; Agent Session Streaming in Preview
GitHub
tools
-
Claude Code v2.1.205: Transcript-Tampering Guard and Session Reliability Fixes
Anthropic
tools
-
Claude Code v2.1.207: Auto Mode GA on Cloud Providers, Security Fix
Anthropic
tools
-
GitHub Copilot Adds Prompt-Injection Detection in CodeQL 2.26.0
GitHub
tools
-
Claude Code v2.1.210–211: Bidirectional-Override Security Patch and Subagent Stream Flag
Anthropic
tools
-
Claude Code v2.1.214–v2.1.215: permission-bypass fixes, EndConversation tool, /verify no longer auto-runs
Anthropic
tools
-
Safeguards Based on Copyable Context Cannot Provide Reliable Safety for LLMs
research
-
Claude Code v2.1.221 adds Focus view and sandbox credential masking
Anthropic
tools
-
Claude Code v2.1.223 fixes permission-bypass gap, merges /code-review
Anthropic
tools
-
Stealing Reasoning Traces from Proprietary LLM APIs
research
-
Claude Code v2.1.232: subagent forking on by default, cross-session mentions, GitLab plugin support
Anthropic
tools
-
OpenClaw 2026.8.1-beta.2 pre-release ships GPT-5.6 Ultra switching and secret-egress binding
tools
-
Claude Code v2.1.235 adds optional spellcheck, fixes window-restore focus and embedded grep
Anthropic
tools
-
Zed v1.16.2 — filesystem sandbox escape fix, GitHub Copilot Chat on GitHub Enterprise Cloud
Zed
tools
-
OpenAI Codex CLI 0.152.0 ships vim search, MCP output limits, credential-protection fix
OpenAI
tools
-
Meta details Muse agent safety architecture and opens its bug bounty to everyone
Meta AI
tools
-
Claude Code v2.1.268 fixes HTTP 400 regression and symlink permission bypass
Anthropic
tools
-
OpenClaw ships 2026.6.35, the final June 2026 LTS release
tools
-
Cline Desktop v0.0.29 patches undici CVE-2026-1525 and enables web search by default
Cline
tools
-
llama.cpp patches remotely exploitable use-after-free in llama-server RPC endpoint
ggml
tools
-
Claude Code v2.1.275/276: claude.ai skill sync, ctrl+enter send-now, npm plugin hardening
Anthropic
tools
-
Codex CLI 0.155.0 goes stable with /voice, Touch ID MCP approvals, sandbox hardening
OpenAI
tools
-
Cline ships SSH remotes and concurrent sub-agents (Desktop 0.0.31/0.0.32, ext 4.1.19)
Cline
tools
-
OpenAI Ships codex-zsh v0.1.0: Versioned Patched zsh Binary for Codex Sandbox
OpenAI
tools
-
OpenAI Codex CLI 0.142.5: Trace-Log Privacy Fix
OpenAI
tools
-
Claude Code v2.1.222 fixes worktree destructive-git-command and hook-bypass safety gaps
Anthropic
tools
-
OpenAI Codex CLI 0.146.1 hardens auto-review for cyber-capable models
OpenAI
tools
-
Claude-Red: a curated offensive-security skills library gains 6k stars
SnailSploit
tools