註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Chris
@ChrisGPT
Agi 2029 - AI Insider / Reporter as featured in Axios • The Information • NYT • Techcrunch
加入 January 2023
1.6K 正在關注    69.1K 粉絲
Wait? How the hell did Tencent compress Hy4 preview from 1.5TB down to 200GB and barely move the benchmarks?? They basically let calibration data decide how aggressively each layer can be compressed. Some layers got pushed all the way down to 1.31 bit, while the more sensitive ones stay closer to 2 bit so the model doesn’t fall apart. Literally insane efficiency gains. And somehow MCP Atlas only moves down to 83.2, and SWE-Bench Multi - 82.9 to 81.3. The low-bit inference progress happening right now is kind of insane.
顯示更多
0
30
963
65
轉發到社區