We've released a full technical report on Prime Agent. Extending from our blog post, we center our discussion around how harnesses should be designed and evaluated. We innovate on 4 fronts:
1. Agentic context management
2. Swarms and depth-n+ RLMs
3. Verifiers support for standardized evals
4. Out-of-loop experiments during autoresearch