• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    Databricks Claims Top Spot for Kimi K3 Inference Speed

    Databricks engineer states the company leads Artificial Analysis for Kimi K3 inference speed.

    GN
    JF
    MZ
    6 Sources, 58d ago, first seen 58d ago

    TLDR

    Yuchen Jin, a Databricks engineer, posted that the company leads Artificial Analysis rankings for Kimi K3 inference speed and latency at 239 tokens per second. He noted the model has 2.8 trillion parameters and is the largest open-source model Databricks has served. Matei Zaharia retweeted the post. Jonathan Frankle retweeted a separate post by Ali Ghodsi that linked to the Artificial Analysis page and stated Databricks has the fastest and lowest latency on Kimi K3. Other replies discussed the result but added no new performance data.

    Combined views

    93K

    6 Sources, first seen 58d ago

    Combined views

    93K

    6 Sources, first seen 58d ago

    1.1K likes
    1.1K likes
    42 comments
    173 saves
    160 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Featured Source
    42 comments
    173 saves
    160 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    6 Sources

    @Yuchenj_UWKimi K3 at 239 tokens/s! Databricks is now #1 for Kimi K3 inference speed and latency on Artificial Analysis. A huge 2.8T parameters model, it’s the largest oss model we’ve ever served. We make sure the GPUs go brrr at Databricks.
    @altryne@Yuchenj_UW Daym, let's GO! Any learnings you guys are willing to share with the open source community?
    @jefrankleRT @alighodsi: @databricks has the fastest and lowest latency on Kimi K3 (max)! Great job by Databricks AI team! https://artificialanalysis.ai/models/kimi-k3/providers h…
    @matei_zahariaRT @Yuchenj_UW: Kimi K3 at 239 tokens/s! Databricks is now #1 for Kimi K3 inference speed and latency on Artificial Analysis. A huge 2.8T…
    @gneubig@Yuchenj_UW That's $13 per hour, not bad for a competent programmer!
    @peteskomorochRT @Yuchenj_UW: Kimi K3 at 239 tokens/s! Databricks is now #1 for Kimi K3 inference speed and latency on Artificial Analysis. A huge 2.8T…

    6 Sources

    @Yuchenj_UWKimi K3 at 239 tokens/s! Databricks is now #1 for Kimi K3 inference speed and latency on Artificial Analysis. A huge 2.8T parameters model, it’s the largest oss model we’ve ever served. We make sure the GPUs go brrr at Databricks.
    @altryne@Yuchenj_UW Daym, let's GO! Any learnings you guys are willing to share with the open source community?
    @jefrankleRT @alighodsi: @databricks has the fastest and lowest latency on Kimi K3 (max)! Great job by Databricks AI team! https://artificialanalysis.ai/models/kimi-k3/providers h…
    @matei_zahariaRT @Yuchenj_UW: Kimi K3 at 239 tokens/s! Databricks is now #1 for Kimi K3 inference speed and latency on Artificial Analysis. A huge 2.8T…
    @gneubig@Yuchenj_UW That's $13 per hour, not bad for a competent programmer!
    @peteskomorochRT @Yuchenj_UW: Kimi K3 at 239 tokens/s! Databricks is now #1 for Kimi K3 inference speed and latency on Artificial Analysis. A huge 2.8T…