• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    Evolution Strategies for LLM Post-Training Examined

    A retweet shares research on using Evolution Strategies to refine large language models without gradients.

    YW
    1 Source, 29d ago, first seen 29d ago

    TLDR

    Yee Whye Teh retweeted a post by @zhengzhi20 that presents a study on Evolution Strategies for LLM reasoning. The post states that ES enables post-training of large language models without backpropagation. It claims the work finds Evolution Strategies achieve higher Pass@K than GRPO during LLM post-training. The accompanying summary describes ES as a gradient-free method that maintains broader exploration compared with other approaches. The discussion centers on this alternative training technique and its reported results for model refinement.

    Combined views

    14

    1 Source, first seen 29d ago

    Combined views

    14

    1 Source, first seen 29d ago

    17 reposts
    17 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    1 Source

    @yeewhyeRT @zhengzhi20: 🧬Fully Understand Evolution Strategies (ES) for LLM Reasoning ES can post-train LLMs without backprop. Our study finds that…

    1 Source

    @yeewhyeRT @zhengzhi20: 🧬Fully Understand Evolution Strategies (ES) for LLM Reasoning ES can post-train LLMs without backprop. Our study finds that…