• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    V4.1 ranks fifth on LiveBench, first in agentic coding, a user reports

    The user credits Python performance for the coding lead, while reporting a TypeScript tie with Fable 5.1 and Muse Spark 1.3 at 60%.

    T(
    4 Sources, 19d ago, first seen 19d ago

    TLDR

    A user discussing LiveBench results on September 11 places V4.1 fifth overall and first in the Agentic Coding category. They attribute the coding lead to Python performance. Their language-by-language breakdown puts V4.1 level with Fable 5.1 and Muse Spark 1.3 in TypeScript at 60%, with a JavaScript score of 81.8 versus 77.3 for each of the next four models.

    Combined views

    21.5K

    4 Sources, first seen 19d ago

    Combined views

    21.5K

    4 Sources, first seen 19d ago

    236 likes
    236 likes
    9 comments
    45 saves
    5 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    9 comments
    45 saves
    5 reposts
    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    4 Sources

    @teortaxesTexIn turn, V4.1's Agentic Coding W is due to Python Supremacy. In TypeScript it's tied with Fable 5.1 and Muse Spark 1.3 at 60%, which I guess means TS score is saturated. In JS it does 81.8, the next 4 models are all 77.3. Looks like spicy slop.

    4 Sources

    @teortaxesTexIn turn, V4.1's Agentic Coding W is due to Python Supremacy. In TypeScript it's tied with Fable 5.1 and Muse Spark 1.3 at 60%, which I guess means TS score is saturated. In JS it does 81.8, the next 4 models are all 77.3. Looks like spicy slop.