• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI
    Announcement

    lithos-metal is being open-sourced, with a claimed peak of 200-plus tokens per second per user on one Apple M5 Max

    The developer says megakernels and DSpark speculative decoding run Qwen3.8-27B at that peak, and shares the code.

    Shreya ShankarSS
    Beidi ChenBC
    Zhihao JiaZJ
    3 Sources, ,

    TLDR

    The lithos-metal developer says it is open-sourcing the code and claims megakernels plus DSpark speculative decoding run Qwen3.8-27B at a peak of 200-plus tokens per second per user on one Apple M5 Max. The developer also says it can be tried with any coding agent in one command.

    Combined views

    11.4K

    3 Sources, first seen 3h ago

    Combined views

    11.4K

    3 Sources, first seen 3h ago

    195 likes
    3h ago
    first seen 3h ago
    195 likes
    12 comments
    114 saves
    33 reposts
    12 comments
    114 saves
    33 reposts

    3 Sources

    Zhihao Jia@JiaZhihaoWe’re open-sourcing lithos-metal 🚀 Megakernels + DSpark speculative decoding run Qwen3.8-27B at 200+ tokens/s/user peak on one @Apple M5 Max. Ultra-fast inference on your laptop. Try it with any coding agent in one command. Code: https://github.com/lithos-ai/lithos-metal Tech blog: https://www.lithosai.com/blog/lithos-metal3h
    Shreya Shankar@sh_reyaRT @JiaZhihao: We’re open-sourcing lithos-metal 🚀 Megakernels + DSpark speculative decoding run Qwen3.8-27B at 200+ tokens/s/user peak on…3h
    Beidi Chen@BeidiChen🤩2h
    Featured Source

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    3 Sources

    Zhihao Jia@JiaZhihaoWe’re open-sourcing lithos-metal 🚀 Megakernels + DSpark speculative decoding run Qwen3.8-27B at 200+ tokens/s/user peak on one @Apple M5 Max. Ultra-fast inference on your laptop. Try it with any coding agent in one command. Code: https://github.com/lithos-ai/lithos-metal Tech blog: https://www.lithosai.com/blog/lithos-metal3h
    Shreya Shankar@sh_reyaRT @JiaZhihao: We’re open-sourcing lithos-metal 🚀 Megakernels + DSpark speculative decoding run Qwen3.8-27B at 200+ tokens/s/user peak on…3h
    Beidi Chen@BeidiChen🤩2h