注册并分享邀请链接,可获得视频播放与邀请奖励。

Macrocosmos
@MacrocosmosAI
Building distributed intelligence on Bittensor. SN1, 9, 13
加入 March 2024
22 正在关注    7.6K 粉丝
Training a 100B parameter model normally means renting a block of matched, high-spec hardware and holding it unbroken for the length of the run. Orion-100B ran across five data centres instead, on single A100s at around $1.25 an hour each. That puts a full replica at roughly $20 an hour, and the entry cost about 2.5x below approaches that need high-spec nodes throughout. It ran as a viability proof rather than to a finished model, and it held. We're in Montreal on Monday talking about what that changes. @ExploitSummit
显示更多