• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    FT Reports ByteDance Pre-Training 10T Parameter Model

    Posts cite Financial Times report on ByteDance pre-training a large AI model.

    SZ
    CP
    RS
    24 Sources, 55d ago, first seen 55d ago

    TLDR

    X users quote a Financial Times article claiming ByteDance is pre-training an AI model with up to 10 trillion parameters. The size is said to be three times larger than Moonshot’s Kimi K3 at 2.8 trillion parameters and close to Anthropic’s Mythos system. Several posts note the reported figure rose from an earlier 5 trillion parameter rumor. The linked FT piece is presented as the source for the details shared in the thread.

    Combined views

    1M

    24 Sources, first seen 55d ago

    Combined views

    1M

    24 Sources, first seen 55d ago

    5.3K likes
    5.3K likes
    359 comments
    724 saves
    365 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    359 comments
    724 saves
    365 reposts
    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    24 Sources

    @scaling01ByteDance is going for a 5T model (im sorry I included google in my AGI tier list half a year go. i was blinded by the hopium, and had longer timelines before Mythos)
    @zephyr_z9Not just ByteDance ;)
    @jukan05FT: BYTEDANCE IS IN THE PRE-TRAINING STAGE OF A MODEL WITH UP TO 10 TRILLION PARAMETERS, APPROACHING MYTHOS IN SCALE.
    @teortaxesTex@zephyr_z9 you could also do like Baidu, REAP + distill lol
    @Scobleizer@zephyr_z9 You are right. Alibaba and others are working on it.
    @suchenzangregistering a prediction that this is going to be a flop the one weakness that particularly competent vp had was not knowing (nor being interested in knowing) any of the details in pretraining i hope i'm wrong!
    @AndrewCurran_RT @jukan05: FT: BYTEDANCE IS IN THE PRE-TRAINING STAGE OF A MODEL WITH UP TO 10 TRILLION PARAMETERS, APPROACHING MYTHOS IN SCALE. https://…
    @sundeep10T 👀👀👀
    @chamathAnd if reports are accurate, with zero distillation to jumpstart it. Thus proving many things that are more value destructive than value accreting…
    @kimmonismusByteDance is training an AI model that could approach the size of Anthropic’s Mythos system Its training a model with as many as 10tn parameters, three times larger than Moonshot’s Kimi K3. 1. Because several reputable sources now confirm it, it can be assumed that it's true and that they are indeed training a 10b parameter model. 2. This makes it even less likely that the US labs will slow down.

    24 Sources

    @scaling01ByteDance is going for a 5T model (im sorry I included google in my AGI tier list half a year go. i was blinded by the hopium, and had longer timelines before Mythos)
    @zephyr_z9Not just ByteDance ;)
    @jukan05FT: BYTEDANCE IS IN THE PRE-TRAINING STAGE OF A MODEL WITH UP TO 10 TRILLION PARAMETERS, APPROACHING MYTHOS IN SCALE.
    @teortaxesTex@zephyr_z9 you could also do like Baidu, REAP + distill lol
    @Scobleizer@zephyr_z9 You are right. Alibaba and others are working on it.
    @suchenzangregistering a prediction that this is going to be a flop the one weakness that particularly competent vp had was not knowing (nor being interested in knowing) any of the details in pretraining i hope i'm wrong!
    @AndrewCurran_RT @jukan05: FT: BYTEDANCE IS IN THE PRE-TRAINING STAGE OF A MODEL WITH UP TO 10 TRILLION PARAMETERS, APPROACHING MYTHOS IN SCALE. https://…
    @sundeep10T 👀👀👀
    @chamathAnd if reports are accurate, with zero distillation to jumpstart it. Thus proving many things that are more value destructive than value accreting…
    @kimmonismusByteDance is training an AI model that could approach the size of Anthropic’s Mythos system Its training a model with as many as 10tn parameters, three times larger than Moonshot’s Kimi K3. 1. Because several reputable sources now confirm it, it can be assumed that it's true and that they are indeed training a 10b parameter model. 2. This makes it even less likely that the US labs will slow down.