• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    Negative Self-Distillation aims to train LLMs to avoid flaws

    An introductory post describes NSD as a label-free method for training large language models to explicitly avoid flaws.

    KK
    1 Source, 15d ago, first seen 15d ago

    TLDR

    A post introducing Negative Self-Distillation (NSD) asks whether large language models could learn to reason by learning what not to do, rather than copying perfect solutions. It presents NSD as a label-free training method built around avoiding flaws.

    Combined views

    1

    1 Source, first seen 15d ago

    reposts

    Combined views

    1

    1 Source, first seen 15d ago

    19 reposts
    19

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    1 Source

    @kastnerkyleRT @peirongcan: 🤔 What if LLMs learn to reason not by copying perfect solutions, but by learning what NOT to do? ❌ 💡Introducing Negative Se…

    1 Source

    @kastnerkyleRT @peirongcan: 🤔 What if LLMs learn to reason not by copying perfect solutions, but by learning what NOT to do? ❌ 💡Introducing Negative Se…