• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI
    Announcement

    At 524k width, BeFOND is claimed to match Gemma Scope’s single-feature probing accuracy with over 1,000 times fewer samples

    A co-lead describes BeFOND as an encoder-free iterative sparse coding model for interpreting language models.

    Dileep GeorgeDG
    Hadi VafaiiHV
    3 Sources, ,

    TLDR

    A BeFOND co-lead says the model is designed to help interpret language models. At a width of 524k, they claim it matches Gemma Scope’s single-feature probing accuracy with over 1,000 times fewer samples—roughly 6 million versus 8 billion SAE-training tokens. They also say BeFOND keeps improving as SAE width increases while alternatives plateau, and describe its inference and learning rules as closed-form.

    Combined views

    1.8K

    3 Sources, first seen 2h ago

    Combined views

    1.8K

    3 Sources, first seen 2h ago

    21 likes
    2h ago
    first seen 2h ago
    21 likes
    2 comments
    14 saves
    4 reposts
    2 comments
    14 saves
    4 reposts
    Featured Source

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    3 Sources

    Hadi Vafaii@hadivafaiiLLMs have shown us that sand can think. But how? We need powerful methods to interpret their mind. ⚫️ Introducing BeFOND, an encoder-free iterative sparse coding model. ⚫️ BeFOND has closed-form inference and learning rules, both unified as natural gradient flow on free energy. ⚫️ At 524k width, BeFOND matches Gemma Scope's single-feature probing accuracy with over 1000x fewer samples (~6M vs 8B SAE-training tokens). ⚫️ BeFOND keeps improving with SAE width, while alternatives plateau. ⚫️ Importantly, we provide theory to explain WHY BeFOND works so well. Read on! 👉🧵[1/17] Work with my amazing co-lead @tejasraoai2h
    Dileep George@dileeplearningRT @hadivafaii: LLMs have shown us that sand can think. But how? We need powerful methods to interpret their mind. ⚫️ Introducing BeFOND,…2h

    3 Sources

    Hadi Vafaii@hadivafaiiLLMs have shown us that sand can think. But how? We need powerful methods to interpret their mind. ⚫️ Introducing BeFOND, an encoder-free iterative sparse coding model. ⚫️ BeFOND has closed-form inference and learning rules, both unified as natural gradient flow on free energy. ⚫️ At 524k width, BeFOND matches Gemma Scope's single-feature probing accuracy with over 1000x fewer samples (~6M vs 8B SAE-training tokens). ⚫️ BeFOND keeps improving with SAE width, while alternatives plateau. ⚫️ Importantly, we provide theory to explain WHY BeFOND works so well. Read on! 👉🧵[1/17] Work with my amazing co-lead @tejasraoai2h
    Dileep George@dileeplearningRT @hadivafaii: LLMs have shown us that sand can think. But how? We need powerful methods to interpret their mind. ⚫️ Introducing BeFOND,…2h