Reactions from ranked influencers
4 postsK3 and other open source models are obviously insanely bullish for compute and RAM. Inference was always going to be almost all high performance compute usage. More people that can do inference = more demand.
Jevons paradox is in full swing already. Cheaper intelligence will create more demand for GPUs. Tokens get cheaper, agents get busier, clusters stay full. That’s the whole loop.
A lot of people make the mistake of thinking that when AI costs drop, that spend on AI drops with it. Usually the opposite happens. When you make AI cheaper, it will get consumed more. Because now you can afford to use for a wider of tasks than you did before. You start to write more code. You review more code for bugs and security. You run agents on large data sets you couldn’t process before. And so on. Thus, for the foreseeable future, anything that lowers the cost of tokens will drive up inference demand. This also gives you some insight also into why even open source business models work in AI. No one is running these models on their devices; they’re running them in infra. Great time to be one of those providers.
the more companies running their own models, the more infra will be needed. Kim3 and open source ecosystem obviously drive demand for compute surprising how the market still sees the fate of the few big lab models as the proxy for the full stack.
Amjad is the founder and CEO of Replit, the best “vibe coding” platform. https://twitter.com/amasad/status/2077989946565206267
A lot of people make the mistake of thinking that when AI costs drop, that spend on AI drops with it. Usually the opposite happens. When you make AI cheaper, it will get consumed more. Because now you can afford to use for a wider of tasks than you did before. You start to write more code. You review more code for bugs and security. You run agents on large data sets you couldn’t process before. And so on. Thus, for the foreseeable future, anything that lowers the cost of tokens will drive up inference demand. This also gives you some insight also into why even open source business models work in AI. No one is running these models on their devices; they’re running them in infra. Great time to be one of those providers.
Kimi K3 has received far more love than we expected, and our GPUs are feeling it. Over the past 48 hours, demand has pushed close to the limits of our current capacity. To protect the experience of existing subscribers, we're temporarily pausing new subscriptions and prioritizing compute for current members. Existing subscribed users are not affected. We're adding capacity as fast as we can and will reopen new subscription spots in batches. Going forward, we'll also split membership into two more focused plans: Kimi Membership for Kimi Web, App, and Work; and Kimi Code Membership for coding workflows. This will help us match compute more precisely and keep the experience stable. Thank you for your patience and understanding!
Combined views
111K
4 posts, first seen 1d ago