• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    Negative Self-Distillation aims to teach LLMs what not to do

    A post introduces Negative Self-Distillation (NSD) as a label-free method for training large language models to explicitly avoid flaws.

    RP
    1 Source, 16d ago, first seen 16d ago

    TLDR

    The announcement asks whether LLMs could learn to reason by learning what not to do, rather than copying perfect solutions. It presents Negative Self-Distillation (NSD) as a label-free training method focused on avoiding flaws.

    Combined views

    16.9K

    1 Source, first seen 16d ago

    140 likes

    Combined views

    16.9K

    1 Source, first seen 16d ago

    140 likes
    1 comments
    155 saves
    23 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    1 comments
    155 saves
    23 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    1 Source

    @peirongcan🤔 What if LLMs learn to reason not by copying perfect solutions, but by learning what NOT to do? ❌ 💡Introducing Negative Self-Distillation (NSD) -- a label-free method that trains LLMs to explicitly avoid flaws! 🧵[0/n]

    1 Source

    @peirongcan🤔 What if LLMs learn to reason not by copying perfect solutions, but by learning what NOT to do? ❌ 💡Introducing Negative Self-Distillation (NSD) -- a label-free method that trains LLMs to explicitly avoid flaws! 🧵[0/n]