• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI
    Announcement

    ExploreNet introduced for adaptive exploration in diffusion RL

    Its creators say ExploreNet adjusts its exploration distribution to the current state, with rewards tied to rollout diversity.

    Stella Li @ COLMSL
    9 Sources, 3h ago, first seen 3h ago

    TLDR

    ExploreNet's creators say its state-conditioned exploration distribution is rewarded for diverse rollouts, aiming for faster, more targeted diffusion GRPO learning. They report that its noise distribution was on average 1.4 times as large as FlowGRPO's; matching the magnitude with isotropic noise suggested targeted exploration was what helped. They also say ExploreNet works with smaller groups and fewer denoising steps, and that annotators perceived greater visual changes when channels it amplified were perturbed.

    Combined views

    4.1K

    9 Sources, first seen 3h ago

    Combined views

    4.1K

    9 Sources, first seen 3h ago

    83 likes
    83 likes
    11 comments
    22 saves
    11 reposts
    Featured Source
    11 comments
    22 saves
    11 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    9 Sources

    Stella Li @ COLM@StellaLisyWe introduce ExploreNet🔍 for learnable exploration in diffusion RL. ExploreNet learns an adaptive exploration distribution conditioned on the current state, rewarded by the diversity of the rollouts. ➡️ Faster, more targeted learning in diffusion GRPO.3h

    9 Sources

    Stella Li @ COLM@StellaLisyWe introduce ExploreNet🔍 for learnable exploration in diffusion RL. ExploreNet learns an adaptive exploration distribution conditioned on the current state, rewarded by the diversity of the rollouts. ➡️ Faster, more targeted learning in diffusion GRPO.3h