Position Overview
About the role
The RL Velocity team owns the efficiency and reliability of our RL Science stack - the infrastructure, tooling, and systems that let researchers iterate quickly on training runs. As a Research Engineer on the team, you'll build and improve the core platform that underpins how we do RL at Anthropic, removing bottlenecks that slow down research and making it easier for the broader org to ship better models faster.
Responsibilities
- Build and improve the RL training infrastructure that researchers depend on day-to-day
- Identify and remove bottlenecks across the RL stack: debugging, profiling, and rearchitecting where needed
- Partner closely with researchers and with adjacent engineering teams (inference, sandboxing, and many more) to understand pain points and ship tooling that makes them faster
- Own the reliability and performance of research runs end-to-end
- Contribute to design decisions that shap...