• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
AI

Nvidia Rubin vLLM inference claimed to have 3.2x better profit per gigawatt than GB300 NVL72

SemiAnalysis also claims up to 10x better performance per dollar than GB300 NVL72 on the vLLM engine.

1 Source, 43m ago, first seen 43m ago

TLDR

SemiAnalysis claims Nvidia Rubin inference running on vLLM offers 3.2x better profit per gigawatt and up to 10x better performance per dollar than GB300 NVL72. It describes vLLM as a widely used production LLM engine.

Combined views

—

1 Source, first seen 43m ago

— likes— comments— saves— reposts

Combined views

—

1 Source, first seen 43m ago

— likes— comments— saves— reposts

Sentiment

Positive——Negative

Summary

Not enough discussion yet.

No sentiment analysis available yet.

Featured Source

Sentiment

Positive——Negative

Summary

Not enough discussion yet.

No sentiment analysis available yet.

1 Source

SemiAnalysis@SemiAnalysis_HOLY ALERT🚨: NVIDIA vLLM INFERENCE RUBIN IS 3.2x BETTER IN PROFIT💰️ PER GIGAWATT & HAS UP TO 🚀 10x BETTER PERF PER DOLLAR THAN EVEN GB300 NVL72. This is on the widely used production LLM engine called @vllm_project. We explain below.👇️(1/3)🧵43m
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI

    1 Source

    SemiAnalysis@SemiAnalysis_HOLY ALERT🚨: NVIDIA vLLM INFERENCE RUBIN IS 3.2x BETTER IN PROFIT💰️ PER GIGAWATT & HAS UP TO 🚀 10x BETTER PERF PER DOLLAR THAN EVEN GB300 NVL72. This is on the widely used production LLM engine called @vllm_project. We explain below.👇️(1/3)🧵43m
    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet