• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    Teortaxes Describes Working RSI as RL Environment Building

    Pseudonymous DeepSeek booster states real RSI involves RL environments and debugging.

    T(
    1 Source, 27d ago, first seen 27d ago

    TLDR

    Teortaxes, a pseudonymous AI developer known for promoting DeepSeek, posted that current functional RSI consists of building and debugging RL environments rather than other approaches. The post notes surprise that this disaggregated mode received no mention in LessWrong discussions. It includes an attached image and references another account. The statement addresses what the author sees as actual progress in recursive self-improvement today.

    Combined views

    5.7K

    1 Source, first seen 27d ago

    Combined views

    5.7K

    1 Source, first seen 27d ago

    29 likes
    29 likes
    1 comments
    25 saves

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    1 comments
    25 saves

    1 Source

    @teortaxesTexFor now, "RSI" that really is happening and reliably working is less like this and more like "build and debug RL environments, do more RL there". In retrospect, strange that this disaggregated mode wasn't in the LessWrong lore.

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    1 Source

    @teortaxesTexFor now, "RSI" that really is happening and reliably working is less like this and more like "build and debug RL environments, do more RL there". In retrospect, strange that this disaggregated mode wasn't in the LessWrong lore.