가입 후 초대 링크를 공유하면 동영상 재생 및 초대 보상을 받을 수 있습니다.

Fireworks
@FireworksAI_HQ
The frontier platform for training and inference on open-weights models at scale.
가입 September 2022
280 팔로잉 중    29.4K
Long-context sparse attention has a catch: data-dependent block selection wrecks memory access kills speed. Our @MiniMax_AI M3 kernel on Blackwell answers it. KV-stationary, each block read once, ~980 TFLOP/s on a B200. See the breakdown here →
더 보기