Position Overview
We're hiring an engineer to help us bring reinforcement learning to every agent team at NVIDIA. This is a rare chance to shape how autonomous, self-improving agents learn and evolve across the enterprise. The role sits at the intersection of ML research and production engineering. What if every agent developer could add self-improvement loops to their workflows without needing deep RL expertise? That's the challenge here: evaluate emerging approaches, adapt them into enterprise-ready blueprints, and make them available inside sandboxed execution environments with the security and governance the enterprise demands. We believe the best training and self-evolving agent platforms come from people with diverse backgrounds and want this person to help us build ours.
What you'll be doing:
The work splits between creating enterprise-ready RL capabilities and partnering with agent teams to put them into practice.
Building RL cookbooks and environments:
+ Eval...