注册并分享邀请链接,可获得视频播放与邀请奖励。

DailyPapers
@HuggingPapers
Tweeting interesting papers submitted at Submit your own at and link models/datasets/demos to it!
加入 March 2025
4 正在关注    21.6K 粉丝
Why the recurrent half of a hybrid LLM is actually easy to quantize Minima quantized all 496 linear layers of Qwen3.8-27B — including Gated DeltaNet — to NVFP4 W4A4, matching BF16 performance while cutting size ~2.9×
显示更多
0
4
160
27
转发到社区