Register and share your invite link to earn from video plays and referrals.

Macrocosmos
@MacrocosmosAI
Building distributed intelligence on Bittensor. SN1, 9, 13
Joined March 2024
22 Following    7.6K Followers
Training a 100B parameter model normally means renting a block of matched, high-spec hardware and holding it unbroken for the length of the run. Orion-100B ran across five data centres instead, on single A100s at around $1.25 an hour each. That puts a full replica at roughly $20 an hour, and the entry cost about 2.5x below approaches that need high-spec nodes throughout. It ran as a viability proof rather than to a finished model, and it held. We're in Montreal on Monday talking about what that changes. @ExploitSummit
Show more