Anthropic's Project Pilot tests whether AI models can fly drones

Anthropic

Research official 1 src. ~1 min

Anthropic's Frontier Red Team, working with Andon Labs, published Project Pilot and a new Drone-Bench benchmark evaluating whether AI models can autonomously fly a quad-rotor drone indoors to locate and follow a person. Fifteen models from three developers were scored on five sub-tasks (reconstruct, localize, navigate, detect, follow); Claude Fable 5 led but reconstruction failures still blocked reliable end-to-end autonomous navigation.

Why it matters

The research documents how close AI models are to autonomously operating physical hardware for surveillance-like tasks, a dual-use capability with direct implications for AI governance and physical-world safety.

Importance: 3/5

Notable frontier-lab safety research release with direct AI-governance implications.

Sources