Agensh: Scaling Organizational Intelligence to 1,024 Agents
Agensh lets up to 1,024 self-organized agents raise ProgramBench scores without a central orchestrator.
Agensh is a self-organized multi-agent harness that avoids a central orchestrator. Workers concurrently gather context, claim sub-tasks, act, share findings, verify results, and merge progress through a shared workspace, message interface, and shared context. On the five hardest ProgramBench tasks with GPT-5.6-sol (high), scaling from 1 to 128 agents raised the mean final test-pass rate from 19.31% to 28.78%. On pandoc, scaling from 1 to 1,024 agents raised the final test-pass rate from 33.89% to 55.06%.