-
OpenAI Previews GPT-5.6 Family: Sol, Terra, and Luna in Government-Gated Limited Release
OpenAI
models-llm
-
Mistral Launches Medium 3.5 Open-Weight Flagship and Remote Coding Agents in Vibe
Mistral AI
models-llm
-
xAI Launches Grok Build: Agentic Coding CLI in Early Beta
xAI
tools
-
Gemini 3.5 Flash Released at Google I/O 2026: Frontier Coding + Agentic at Flash Speed
Google DeepMind
models-llm
-
Alibaba Launches Qwen3.7-Plus: Multimodal Agent with Vision, Reasoning, and Autonomous Execution
Alibaba / Qwen
models-llm
-
MiniMax Releases M3: Open-Weight Frontier Model with 1M-Token Context and MSA Architecture
MiniMax
models-llm
-
MiniMax M3 Open Weights Released: 1M Context, MoE, Frontier Coding
MiniMax
models-llm
-
xAI Launches Grok 4.5: Cursor-Trained Coding Model with 2x Token Efficiency
xAI
models-llm
-
Tencent Releases Hunyuan Hy3: 295B Open-Weight MoE Model Under Apache 2.0
Tencent
models-llm
-
OpenAI Launches ChatGPT Work and Unifies Codex Into Desktop App
OpenAI
tools
-
Moonshot AI launches Kimi K3, a 2.8T-parameter open-weight model
Moonshot AI
models-llm
-
Anthropic releases Claude Opus 5
Anthropic
models-llm
-
Google Launches Gemini Spark: 24/7 Personal AI Agent in Google AI Ultra
Google
tools
-
Google Launches Antigravity 2.0: Agent-First Dev Platform with Desktop App, CLI, and Managed Agents API
Google
tools
-
Code as Agent Harness: Survey Positions Code as the Substrate for Executable Agent Systems (159 HF upvotes)
Multi-institution (42 authors)
research
-
xAI Launches Composer 2.5 in Grok Build for Agentic Coding
xAI
models-llm
-
OpenAI Expands Codex Beyond Developers: Sites, Annotations, and Six Role-Specific Business Plugins
OpenAI
tools
-
Gemini 3.5 Flash Gains Native Computer Use as Built-in Tool
Google DeepMind
tools
-
Yandex Launches AI Agent Platform for Alice AI
Yandex
tools
-
Gemini Spark Launches on macOS with Local File Access and MCP Server Support
Google
tools
-
LangChain and NVIDIA Launch NemoClaw Deep Agents Blueprint for Enterprise Open Agents
LangChain
tools
-
Google releases Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
Google DeepMind
models-llm
-
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning
NVIDIA
research
-
Claude Code v2.1.219 Sets Claude Opus 5 as Default Opus Model
Anthropic
tools
-
GitHub Copilot Adds Claude Opus 5 for Agentic Coding Tasks
Anthropic / GitHub
tools
-
Moonshot AI Releases Kimi K2.7-Code: 1T-Parameter Open-Weight Coding Model with Vision
Moonshot AI
models-llm
-
Kimi K3 Technical Report: Kimi Delta Attention and Stable LatentMoE Architecture Detailed
Moonshot AI
research
-
SDAR: Self-Distilled Agentic Reinforcement Learning for Multi-Turn Agents
Zhejiang University / Meituan
research
-
Moonshot AI Opens Kimi Work Desktop Agent with 300-Sub-Agent Swarm and WebBridge
Moonshot AI
tools
-
Runway Launches Agent Skills for Autonomous Ad Campaign and Commercial Production
Runway
video
-
Claude Code v2.1.200: Manual Permission Mode Becomes Default
Anthropic
tools
-
CompactionRL: Reinforcement Learning with Context Compaction for Long-Horizon Agents
Zhipu AI / Tsinghua University
research
-
GitHub Copilot desktop app goes GA on all plans including Free, with BYOK support
GitHub
tools
-
AutomationBench-AA: 657-task independent benchmark for AI agent SaaS automation
Artificial Analysis
tools
-
Zhipu AI Releases GLM-5.2: 744B MoE with 1M-Token Context and Coding-First Design
Zhipu AI
models-llm
-
FAPO: Fully Autonomous Prompt Optimization of Multi-Step LLM Pipelines
Cisco Foundation AI
research
-
Qwen-Image-Agent: Agentic Context Building to Bridge the Prompt Underspecification Gap in T2I
Qwen (Alibaba)
research
-
VK Rolls Out Discovery AI Neural Search Across VK Video, Mail Media, and Dzen
VK AI
tools
-
OpenAI Reorganizes Product Teams Around Agentic Strategy, Brockman Takes Charge
OpenAI
industry
-
Grok 4.5 Enters Private Beta at SpaceX and Tesla
xAI
models-llm
-
Playful Agentic Robot Learning: Self-Directed Play Yields Transferable Robot Skills
UC Berkeley
research
-
Cura 1T: a specialized model for agentic healthcare
actAVA AI
research
-
Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills
Alibaba
research
-
Yandex Consolidates AI Teams Under Alice AI with New Leadership Appointments
Yandex
industry