Prime Intellect 
Reinforcement learning & environments
Build training environments for language model agents, with verifiers that check reasoning and tool use. Run post-training experiments, inspect rollouts, and refine reward signals to help models complete longer tasks and recover when a step goes wrong.