Register and share your invite link to earn from video plays and referrals.

filipe
@filicroval
data eng | 1xAsus Ascent GX10 | benchmarking local models so you don't have to
Joined May 2023
214 Following    153.4K Followers
legendary drop from NVIDIA: ModelOpt 0.45.0 biggest additions: - New NVFP4 (W4A16) weight-only quantization format that requires no calibration - Better MoE support (including mixed NVFP4 + FP8 recipes for models like Nemotron) - Easy MXFP4 → NVFP4 conversion for models like DeepSeek V4 and GPT-OSS - Various improvements for large-scale PTQ and Megatron workflows looks like NVIDIA is pushing harder on making 4-bit inference more practical and calibration-free.
Show more