• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI
    Report

    Manager-free coding-agent teams reportedly scored higher as they grew

    A post describing a Microsoft paper says average scores rose at every step from one to 128 agents on the five hardest of ProgramBench's 200 tasks.

    RP
    1 Source, 2h ago, first seen 2h ago

    TLDR

    A post describing a Microsoft paper says larger manager-free coding-agent teams scored higher and got there sooner. On ProgramBench's five hardest tasks, it says average scores rose at every step from one to 128 agents. Agensh agents claim subtasks, build and test work, merge it into a shared Git repo and log findings on a shared board. The post says each run used one model, and the paper did not report large-team costs.

    Combined views

    3.3K

    1 Source, first seen 2h ago

    Combined views

    3.3K

    1 Source, first seen 2h ago

    39 likes
    39 likes
    18 comments
    21 saves
    5 reposts
    18 comments
    21 saves
    5 reposts
    Featured Source

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    1 Source

    @rohanpaul_aiTurns out coding agents don't need a boss: New Microsoft paper finds that bigger teams of coding agents score higher and get there sooner when agents claim their own tasks without a central manager. Even on the 5 hardest of ProgramBench's 200 tasks, scaling a manager-free team from 1 to 128 agents raised the average score at every step. so for big jobs, add agents and let them coordinate through shared tools with no lead agent. Popular multi-agent coding tools send all work through 1 lead agent, which can only manage so many helpers. Agensh drops the lead agent. Each agent claims a sub-task, builds and tests it, merges it into a shared Git repo, and logs findings on a shared board. Every run used 1 model, and the paper does not report what large teams cost.2h
    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet