Software Engineer, RL Training Infra

OpenAI · San Francisco, California, United States

Location

San Francisco

Job Type

Full time

Posted

June 07, 2026

Job Description

About the Team
The Post-Training Frontiers team creates the frontier agents OpenAI ships to the world. We do the reinforcement learning training for the agentic models we ship in Codex, ChatGPT, and the API (from o1 to 5.5).
Our role consists of shepherding all integrations that should go into the final RL run and deciding what can make it in, babysitting and scaling the final run, and building the research and infra for horizontal integrations, such as improving function calling, factuality, multi-agent capabilities, memory, calibrated thinking, etc.
About the Role
This role focuses on keeping our frontier RL training runs fast, reliable, and unblocked. You will work across engineering and infrastructure problems as they emerge, from scaling and orchestration issues to inference bottlenecks, numerical problems, and hardware failures, as well as supporting large horizontal integrations in the big run, like multi-agent capabilities or memory. Thi...
        

Ready to Apply?

Submit your application for Software Engineer, RL Training Infra at OpenAI

Apply Now