• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    Jonas Geiping Announces Recurrent Depth LLM Training

    Machine learning researcher Jonas Geiping announces training of an LLM with recurrent depth.

    JG
    1 Source, 597d ago, first seen 597d ago

    TLDR

    Jonas Geiping, a machine learning researcher, posted that his group trained an LLM with recurrent depth at scale. He described spending the last year, actually a bit longer, on the project. The model includes an internal latent space that lets it adaptively spend more compute to think longer. Geiping said he could finally discuss the work publicly and pointed to a tech report in the post.

    Combined views

    372.1K

    1 Source, first seen 597d ago

    Combined views

    372.1K

    1 Source, first seen 597d ago

    2.1K likes
    2.1K likes
    54 comments
    1.6K saves
    196 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    54 comments
    1.6K saves
    196 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    1 Source

    @jonasgeipingOk, so I can finally talk about this! We spent the last year (actually a bit longer) training an LLM with recurrent depth at scale. The model has an internal latent space in which it can adaptively spend more compute to think longer. I think the tech report ...🐦‍⬛

    1 Source

    @jonasgeipingOk, so I can finally talk about this! We spent the last year (actually a bit longer) training an LLM with recurrent depth at scale. The model has an internal latent space in which it can adaptively spend more compute to think longer. I think the tech report ...🐦‍⬛