• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI
    Announcement

    Copying peers’ discoveries may narrow exploration among self-improving LLM agents

    In controlled tests, the paper’s authors found that three LLMs earned less reward per token than solo learners.

    X(
    RS
    KJ
    3 Sources, ,

    TLDR

    A paper on recursive social improvement examines whether self-improving LLM agents benefit from learning from peers. Its authors found that three LLMs earned less reward per token than solo learners in controlled tests. Observing peers helped one model find useful skills sooner and another spend less on private search, but neither outperformed independent learners at the same cost. Sharing skills also concentrated agents around fewer independent discoveries.

    Combined views

    3.9K

    3 Sources, first seen 11h ago

    Combined views

    3.9K

    3 Sources, first seen 11h ago

    67 likes
    11h ago
    first seen 11h ago
    67 likes
    4 comments
    28 saves
    27 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Featured Source
    4 comments
    28 saves
    27 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    3 Sources

    @kjha02LLM agents are evolving their own skills and working together in swarms. But how well do they learn from each other? Our new paper on Recursive Social Improvement finds that copying peers’ discoveries can come at the expense of exploration! 🧵 https://arxiv.org/abs/2609.3851611h
    @RulinShaoRT @kjha02: LLM agents are evolving their own skills and working together in swarms. But how well do they learn from each other? Our new p…7h
    @xuanalogueRT @kjha02: LLM agents are evolving their own skills and working together in swarms. But how well do they learn from each other? Our new p…27m

    3 Sources

    @kjha02LLM agents are evolving their own skills and working together in swarms. But how well do they learn from each other? Our new paper on Recursive Social Improvement finds that copying peers’ discoveries can come at the expense of exploration! 🧵 https://arxiv.org/abs/2609.3851611h
    @RulinShaoRT @kjha02: LLM agents are evolving their own skills and working together in swarms. But how well do they learn from each other? Our new p…7h
    @xuanalogueRT @kjha02: LLM agents are evolving their own skills and working together in swarms. But how well do they learn from each other? Our new p…27m