JobMesh

Research Engineer, Machine Learning (RL Velocity)

Anthropic · London, England, GB

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and fo...

Job description

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About the role: The RL Velocity team owns the efficiency and reliability of our RL Science stack - the infrastructure, tooling, and systems that let researchers iterate quickly on training runs. As a Research Engineer on the team, you'll build and improve the core platform that underpins how we do RL at Anthropic, removing bottlenecks that slow down research and making it easier for the broader org to ship better models faster. This is high-leverage work: small improvements to velocity compound across every researcher and every run. Responsibilities: You may be a good fit if you - Build and improve the RL training infrastructure that researchers depend on day-to-day - Identify and remove bottlenecks across the RL stack: debugging, profiling, and rearchitecting where needed - Partner closely with researchers and with adjacent engineering teams (infere...