We're hiring for the Code RL team within the RL organization. As a Research Engineer, you'll advance our models' ability to safely write correct, fast code for accelerators.
You'll need to know accelerator performance well to turn it into tasks and signals models can learn from. Specifically, you will:
Invent, design and implement RL environments and evaluations.
Conduct experiments and shape our research roadmap.
Deliver your work into training runs.
Collaborate with other researchers, engineers, and performance engineering specialists across and outside Anthropic.
You may be a good fit if you:
Have expertise with accelerators (CUDA, ROCm, Triton, Pallas), ML framework programming (JAX or PyTorch).
Have worked across the stack – kernels, model code, distributed systems.
Know how to balance research exploration with engineering implementation.
Are passionate about AI's potential and committed to developing safe and beneficial systems.