• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    Engineer Lists Inference Costs for Large AI Models

    Research engineer Florian Brand lists inference prices for four large models of similar scale.

    LA
    FB
    3 Sources, 32d ago, first seen 32d ago

    TLDR

    Florian Brand, a research engineer at Prime Intellect, replied with size and price figures for four models. Nemotron is listed at 550B parameters and 2.40/M. Trinity is listed at 398B and 0.80/M. Qwen MoE is listed at 397B and 3.50/M. GLM Flash is listed at 320B and 0.50/M. Brand added that readers should not focus on one parameter of a model. The post supplies no further verification or context beyond these stated values.

    Combined views

    25.8K

    3 Sources, first seen 32d ago

    Combined views

    25.8K

    3 Sources, first seen 32d ago

    68 likes
    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    68 likes
    3 comments
    5 saves
    1 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    3 comments
    5 saves
    1 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    3 Sources

    @xeophon@scaling01 Nemotron is 550B and costs 2.40/M Trinity is 398B and costs 0.80/M Qwen MoE is 397B and costs 3.50/M GLM Flash is 320B and costs 0.50/M Don’t focus on one param of a model
    @scaling01@xeophon bad comparison. not the pricing from one company within the same month

    3 Sources

    @xeophon@scaling01 Nemotron is 550B and costs 2.40/M Trinity is 398B and costs 0.80/M Qwen MoE is 397B and costs 3.50/M GLM Flash is 320B and costs 0.50/M Don’t focus on one param of a model
    @scaling01@xeophon bad comparison. not the pricing from one company within the same month