A market for the learning frontier. Decentralized GRPO training on Bittensor Subnet 81| every rollout cryptographically verified before it trains the checkpoint
Introducing Reliquary-4B.
A 4B math & code model trained with reinforcement learning. Anyone could join the network and contribute rollouts.
Independent miners chose the prompts and generated the rollouts. The protocol verified them and trained the model.
Here’s the model and the research behind it.