Y Combinator says a better harness lifts the same AI model from 30% to 95% on ARC-AGI
YC says it gathered researchers and founders to explore AI harnesses—the software around models—and lessons from building an agent for every employee at the company.
TLDR
AI harnesses deserve to be treated as research, not dismissed as scaffolding or prompt engineering, Y Combinator argues. It says the same model weights that score 30% on ARC-AGI score 95% with a better harness. YC describes a discussion spanning self-improving harnesses, messaging between agents, personal AI on local devices and its own workplace harness, QM. It also says the discussion covers what YC learned building an agent for every employee.
Combined views
439.6K
8 Sources, first seen 27d ago