註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Benjamin Marie
@bnjmn_marie
Independent AI researcher (LLM, NLP). My blog, The Kaitchup - AI on a Budget:
加入 June 2019
221 正在關注    6.9K 粉絲
"Full" NVFP4 Qwen3.8 27B: Did you try this? I can't trust the eval. LiveCodeBench results for the BF16 baseline are 10~15 points below what they should be. Same for AIME. Evaluated with a low context length (32K), so in a condition where seeing quantization degradation is unlikely Just wondering whether I should spend one or two days on an RTX Pro 6000 to confirm, but at native context length.
顯示更多