How models and their surrounding harnesses factor into AI agent quality
A user highlights a paper by 17 researchers that maps six harness components and matches setups to different tasks.
TLDR
A user describes a paper by 17 researchers, led by the former head of Huawei’s Noah’s Ark AI lab. Its central claim, as the user presents it, is that AI-agent quality depends on both the model and the harness around it, not the model alone. The post points to a six-part harness, task-specific setups and SWE-bench results broken out by model and harness.
Combined views
94
1 Source, first seen ago
30 reposts
