• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    Tool use across OpenAI models in a 616,300-response review

    SemiAnalysis says GPT 5.6 Terra averaged 1.94 client tool calls per response in its internal usage, compared with 1.00 for Sol and 0.97 for Luna.

    SemiAnalysisSE
    2 Sources, 20d ago, first seen 20d ago

    TLDR

    SemiAnalysis examined 616,300 OpenAI responses from its own internal usage. It reports that GPT 5.6 Terra averaged 1.94 client tool calls per response, versus 1.00 for Sol and 0.97 for Luna. Its early GPT 6 Astra usage fell between those rates, at 1.18 calls per response across 11,717 responses.

    Combined views

    36.8K

    2 Sources, first seen 20d ago

    Combined views

    36.8K

    2 Sources, first seen 20d ago

    146 likes
    146 likes
    20 comments
    36 saves
    16 reposts
    20 comments
    36 saves
    16 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    2 Sources

    SemiAnalysis@SemiAnalysis_Astra's higher average is interesting because it used tools in a smaller share of responses than Sol, 93.5% versus 96.7%. But when it did use tools, we found it averaged more calls: 1.27 versus 1.04. As a reminder, these are representative of our specific workloads making it hard to compare with other data. The models tend see different task classifications, Astra's observation window is short, and the earlier GPT cohorts are much smaller than Sol's. Do you see similar patterns in your own workflows? (2/2)20d

    2 Sources

    SemiAnalysis@SemiAnalysis_Astra's higher average is interesting because it used tools in a smaller share of responses than Sol, 93.5% versus 96.7%. But when it did use tools, we found it averaged more calls: 1.27 versus 1.04. As a reminder, these are representative of our specific workloads making it hard to compare with other data. The models tend see different task classifications, Astra's observation window is short, and the earlier GPT cohorts are much smaller than Sol's. Do you see similar patterns in your own workflows? (2/2)20d