NVIDIA Releases NVFP4 Quantized Qwen3.8-Flash-Next on Hugging Face
Automated account shares NVIDIA's release of a smaller 125B MoE model.
TLDR
DailyPapers, an automated curation account that tweets AI and ML papers from Hugging Face, posted that NVIDIA released the NVFP4 quantized Qwen3.8-Flash-Next. The post describes a 125B MoE model with hybrid attention now 63 percent smaller with minimal accuracy loss. It includes a link to an image showing a horizontal flowchart of five sequential stages of green hexagonal neural networks on a black background with green graphics. The account promotes submissions to the platform's papers section.
Combined views
67.5K
1 Source, first seen 26d ago