Anthropic is seeking a Research Engineer within Reinforcement Learning to advance capabilities and safety of large language models. You will implement novel approaches and contribute to research direction while collaborating with researchers and engineers.
You will design and optimize RL infrastructure, create training environments, and push state-of-the-art methods in agentic model development, tool use, and reasoning.
#J-18808-Ljbffr…
