Prime Agent: A Self-Improving RLM Harness

Prime Intellect

Research official + media 3 src. ~1 min

Open-source harness for long-horizon coding and evaluation workflows. Persistent IPython REPL implements the Recursive Language Model abstraction; Continual Harness preserves histories/memories/skills/subagent specs across trajectories; recursive subagents coordinate via direct agent-to-agent messaging. On ARC-AGI-3 RHAE Best@1 the harness lifts a base model from 30% to 95.5%, and matches or beats native and popular harnesses on long-context coding, GPU kernel generation, emulator construction, and autonomous nanoGPT speedruns.

Why it matters

HF Daily paper with 18.2k cumulative upvotes — a single harness choice is enough to more than triple ARC-AGI-3 performance, suggesting many 'model capability' numbers are partly harness artifacts. Demonstrated refinement and parallelized subagents on Factorio.

Importance: 2/5

default

Sources

official arXiv abstract