We're publishing over 365,000 open and agentic RL Environments for SWE, terminal, and search agents
The open research ecosystem has produced many great datasets for the three main agentic domains - software engineering, terminal use, and web research - but every one of them ships with its own harness, its own image conventions, its own grading scripts, and its own failure modes.
We integrated them all.
23 tasksets behind one API, one sandbox lifecycle, one command.
365,000+ tasks in total,
~198,000 software engineering tasks across 20+ languages
~28,600 terminal tasks
~137,600 search tasks
Ready for evals and RL training on Prime Intellect infrastructure, with validated and cleaned dataset re-uploads where the originals needed fixing.