• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    Maithra Raghu Says Fable 5.1 Leads FrontierFinance Scores

    Samaya AI CEO shares partnership evaluation results with Anthropic on new model.

    MR
    1 Source, 29d ago, first seen 29d ago

    TLDR

    Maithra Raghu posted that her team at Samaya AI partnered with AnthropicAI to test ClaudeAI Fable 5.1 before launch on the FrontierFinance benchmark. She stated Fable 5.1 achieved 55.9 percent, ahead of Fable 5 at 49.2 percent, and called it the top frontier model by a wide margin. Raghu noted the improvement stems from the model performing more and better tasks on the evaluation. The post identifies her as co-founder and CEO of Samaya AI, an AI knowledge-discovery platform for finance, and former Google Brain research scientist.

    Combined views

    5.6K

    1 Source, first seen 29d ago

    Combined views

    5.6K

    1 Source, first seen 29d ago

    52 likes
    52 likes
    2 comments
    16 saves
    11 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    2 comments
    16 saves
    11 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    1 Source

    @maithra_raghuWe partnered with @AnthropicAI to evaluate @claudeai Fable 5.1 pre-launch on FrontierFinance! Fable 5.1 becomes the best performing frontier model by a wide margin, scoring 55.9% compared to 49.2% for Fable 5 (the prior best performing frontier model.) This gain largely comes from Fable 5.1 carrying out more and better tool calls, e.g. calling and anchoring better on authoritative sources. The increased tool calls does come with increased cost though -- we observe Fable 5.1 is about 1.7x as expensive as Fable 5. Exciting results and lots more work ahead to continue advancing agentic capabilities!

    1 Source

    @maithra_raghuWe partnered with @AnthropicAI to evaluate @claudeai Fable 5.1 pre-launch on FrontierFinance! Fable 5.1 becomes the best performing frontier model by a wide margin, scoring 55.9% compared to 49.2% for Fable 5 (the prior best performing frontier model.) This gain largely comes from Fable 5.1 carrying out more and better tool calls, e.g. calling and anchoring better on authoritative sources. The increased tool calls does come with increased cost though -- we observe Fable 5.1 is about 1.7x as expensive as Fable 5. Exciting results and lots more work ahead to continue advancing agentic capabilities!