@NVIDIAAI just gifted us a 75B MoE ๐คฉ๐คฉ๐คฉ
nvidia/NVIDIA-Nemotron-Labs-3-Puzzle-75B-A9B-NVFP4
75.3B total / 9.3B active compressed from Nemotron-3-Super-120B using the Iterative Puzzle framework.
1M token context support!
Perfect for your single GB10 โฅ๏ธ