• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    Will Brown Compares Past and Current AI Models

    Replies compare reasoning capabilities of models from a year ago and today.

    T(
    WB
    2 Sources, 27d ago, first seen 27d ago

    TLDR

    In replies discussing future reasoning benchmarks, will brown states that the leading AA model from twelve months prior has now been matched by muse glimmer, described as a current open-source system. The exchange began with teortaxesTex asking about the smallest model expected to reach a particular capability level within the next year, using examples of prior releases. Brown responds directly to that query by naming the historical leader and its recent equivalent. The posts focus on specific model designations and their relative standing in agentic performance without further elaboration on methods or broader implications.

    Combined views

    3.3K

    2 Sources, first seen 27d ago

    Combined views

    3.3K

    2 Sources, first seen 27d ago

    15 likes
    15 likes
    6 comments
    4 saves

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    6 comments
    4 saves

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    2 Sources

    @teortaxesTexConfusing question: Parameter count of the smallest model on Astra level *in reasoning* in 12 mo (leaked or disclosed), according to the formula sqrt(total B * active B) eg: DS-V4-Flash is sqrt(284*13)=60B; V4-Pro is sqrt(1600*49)=280; K3 is 540; "Mythos Preview" 10T 250AB = 1581
    @willcb@teortaxesTex the best AA (v4.1) model from 12mo ago was gpt-5 with a score of 35, matched today by muse glimmer (30b dense oss)

    2 Sources

    @teortaxesTexConfusing question: Parameter count of the smallest model on Astra level *in reasoning* in 12 mo (leaked or disclosed), according to the formula sqrt(total B * active B) eg: DS-V4-Flash is sqrt(284*13)=60B; V4-Pro is sqrt(1600*49)=280; K3 is 540; "Mythos Preview" 10T 250AB = 1581
    @willcb@teortaxesTex the best AA (v4.1) model from 12mo ago was gpt-5 with a score of 35, matched today by muse glimmer (30b dense oss)