• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    DiffusionGemma blog announcement promises a look at what works in post-training

    The author describes DiffusionGemma as one of the first large, open-weight uniform diffusion language models and asks how the community can post-train it.

    AM
    1 Source, 19d ago, first seen 19d ago

    TLDR

    The announcement says a new blog explores post-training DiffusionGemma—what works, what breaks and which supervised fine-tuning (SFT) objective wins.

    Combined views

    9.6K

    1 Source, first seen 19d ago

    101 likes

    Combined views

    9.6K

    1 Source, first seen 19d ago

    101 likes
    3 comments
    58 saves
    18 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    3 comments
    58 saves
    18 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    1 Source

    @andreamiele_DiffusionGemma is one of the first large, open-weight uniform diffusion LLMs. But how can the community actually post-train it? New blog 📁: what works, what breaks, and the SFT objective that wins. 🧵

    1 Source

    @andreamiele_DiffusionGemma is one of the first large, open-weight uniform diffusion LLMs. But how can the community actually post-train it? New blog 📁: what works, what breaks, and the SFT objective that wins. 🧵