Anthropic's Project Pilot tests whether AI models can fly drones
Anthropic
Anthropic's Frontier Red Team, working with Andon Labs, published Project Pilot and a new Drone-Bench benchmark evaluating whether AI models can autonomously fly a quad-rotor drone indoors to locate and follow a person. Fifteen models from three developers were scored on five sub-tasks (reconstruct, localize, navigate, detect, follow); Claude Fable 5 led but reconstruction failures still blocked reliable end-to-end autonomous navigation.
Why it matters
The research documents how close AI models are to autonomously operating physical hardware for surveillance-like tasks, a dual-use capability with direct implications for AI governance and physical-world safety.
Importance: 3/5
Notable frontier-lab safety research release with direct AI-governance implications.