• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    NVIDIA Releases NVFP4 Quantized Qwen3.8-Flash-Next on Hugging Face

    Automated account shares NVIDIA's release of a smaller 125B MoE model.

    DA
    1 Source, 26d ago, first seen 26d ago

    TLDR

    DailyPapers, an automated curation account that tweets AI and ML papers from Hugging Face, posted that NVIDIA released the NVFP4 quantized Qwen3.8-Flash-Next. The post describes a 125B MoE model with hybrid attention now 63 percent smaller with minimal accuracy loss. It includes a link to an image showing a horizontal flowchart of five sequential stages of green hexagonal neural networks on a black background with green graphics. The account promotes submissions to the platform's papers section.

    Combined views

    67.5K

    1 Source, first seen 26d ago

    Combined views

    67.5K

    1 Source, first seen 26d ago

    709 likes
    709 likes
    10 comments
    480 saves
    39 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    10 comments
    480 saves
    39 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    1 Source

    @HuggingPapersNVIDIA just released the NVFP4 quantized Qwen3.8-Flash-Next on Hugging Face 125B MoE with hybrid attention, now 63% smaller with minimal accuracy loss.

    1 Source

    @HuggingPapersNVIDIA just released the NVFP4 quantized Qwen3.8-Flash-Next on Hugging Face 125B MoE with hybrid attention, now 63% smaller with minimal accuracy loss.