Research Engineer, RL Engineering is a AI Research Engineer role (full-time). with Anthropic. in SEATTLE, US. Compensation shown: $500K–$850K. Imported listing (source: aihiringboard.com). Apply on the employer's site (aihiringboard.com).
Imported listing (source: aihiringboard.com) · Apply on aihiringboard.com
The sections below reproduce the third-party job description for reference. AIEngineer.careers does not write or control this text.
Imported job description
Sourced from aihiringboard.com
Anthropic is hiring a Research Engineer for its RL Engineering team, which builds and owns the reinforcement learning training system behind Claude. You will work across the full stack—orchestration, environments, training, inference, and evaluation—to improve how RL training behaves at scale for both production and research runs. The role sits at the center of RL at Anthropic, collaborating closely with research teams across the company.
Responsibilities
Build, own, and improve the core RL training system serving production and research runs
Work across the stack: orchestration, environments, training, inference, and evaluation
Study how RL training behaves at scale and contribute to research that improves it
Implement new training methods as stable, fast, well-tested code
Improve speed and efficiency of RL training and evaluation through profiling, optimization, and benchmarking
Make the system easier for researchers via clean abstractions, clear APIs, and automated testing
Debug hard problems across the stack, including distributed systems failures at scale
Communicate results clearly in writing and discussion
Requirements
Proficiency in Python and experience debugging and improving large ML codebases
Anthropic is a premier AI research and safety company that develops the Claude family of foundation models, focusing on safety, interpretability, and steerability. It operates as a public benefit corporation.