가입 후 초대 링크를 공유하면 동영상 재생 및 초대 보상을 받을 수 있습니다.

filipe
@filicroval
data eng | 1xAsus Ascent GX10 | benchmarking local models so you don't have to
가입 May 2023
214 팔로잉 중    153.4K
legendary drop from NVIDIA: ModelOpt 0.45.0 biggest additions: - New NVFP4 (W4A16) weight-only quantization format that requires no calibration - Better MoE support (including mixed NVFP4 + FP8 recipes for models like Nemotron) - Easy MXFP4 → NVFP4 conversion for models like DeepSeek V4 and GPT-OSS - Various improvements for large-scale PTQ and Megatron workflows looks like NVIDIA is pushing harder on making 4-bit inference more practical and calibration-free.
더 보기