• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    Randomized YaRN Introduced for LLM Reasoning

    Post claims standard YaRN falls short for long-context reasoning tasks.

    GD
    1 Source, 26d ago, first seen 26d ago

    TLDR

    Associate professor Greg Durrett retweeted a post by manasmehta20. The post states that YaRN is not enough for LLM reasoning at 128K context. It introduces Randomized YaRN as an alternative. The method involves training on short-context data with randomized approaches. The announcement comes directly from the author of the post. No independent corroboration or additional outcomes appear in the visible posts. The conversation centers on this specific claim about model performance limits and the proposed training change.

    Combined views

    2

    1 Source, first seen 26d ago

    Combined views

    2

    1 Source, first seen 26d ago

    1 reposts
    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    1 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    1 Source

    @gregd_nlpRT @manasmehta20: YaRN is not enough for LLM reasoning at 128K context. 🧶Introducing Randomized YaRN: train on short-context data with rand…

    1 Source

    @gregd_nlpRT @manasmehta20: YaRN is not enough for LLM reasoning at 128K context. 🧶Introducing Randomized YaRN: train on short-context data with rand…