Register and share your invite link to earn from video plays and referrals.

Macrocosmos
@MacrocosmosAI
Building distributed intelligence on Bittensor. SN1, 9, 13
Joined March 2024
22 Following    7.6K Followers
We trained Orion-16B without a reserved cluster anywhere in it. The numbers underneath that: It sustained around 80k tokens per second across 4090s, 5090s, A100s, L40s, and A6000s, at roughly 20% MFU. The accessible pool came to about 18 B200s equivalent at FP16. Nodes dropped, throughput swung, and machines ran inconsistently throughout. Training carried on without direct intervention. Imperfect compute, one finished model.
Show more