Stanford introduced DeLM, a decentralized approach to multi-agent systems.
DeLM improves accuracy by up to 17.5 percent points and runs up to 2.49× faster than the Codex and Claude Code harnesses on long-horizon tasks from Terminal-Bench 4.0, DeepSWE v1.1, and ProgramBench.
How it works:
the main agent is replaced by a shared context and a task queue. Agents claim tasks asynchronously and build on or correct each other’s progress, and every agent’s status is visible to the others.
You can run DeLM’s collaborating agents directly in Codex and Claude Code
Post #4532
163