註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Jon Durbin
@jon_durbin
Human. Backend dev
加入 December 2012
138 正在關注    7.2K 粉絲
TL;DR: libp2p TCP transport with dual trainer/syncer roles = chef's kiss For model training, don't bottleneck and single-point-of -failure yourself with a blob store (S3/R2). Also, don't bash your forehead against a wall using listening sockets on nodes that may be behind firewalls/NAT/port mapped containers/etc. Just set up a separate backbone (sync only, non-GPU/training nodes) across the world with super fast WAN and use push only from training/GPU nodes to this layer. Easy peasy, works like a charm. And bonus, libp2p's TCP transports are reliably better across providers/networks/countries compared to default quic/udp. Skynet?
顯示更多
0
10
60
10
轉發到社區