← Back to opportunities

Training, Process Management Engineer

📍 Location
London
⏰ Job Type
Full time
📅 Posted
June 06, 2026

About the Role

About the Team

Training Runtime designs the core distributed runtime that powers everything from early research experiments to frontier-scale model runs. We work on building robust, scalable, high performance components to support our distributed training workloads. Our priorities are to maximize the productivity of our researchers and our hardware, with the goal of accelerating progress towards AGI.

Within Training Runtime, the Process Management team develops the distributed OS responsible for launching, coordinating, and supervising the large numbers of processes that make up modern training workloads. Our runtime sits beneath training frameworks and on top of research infrastructure, ensuring jobs run reliably across massive clusters while maintaining performance, stability, and observability.

Success for us is measured by both system reliability and researcher velocity - enabling ideas to scale from experiments to production trainin...

Ready to Join Through a Referral?

Apply now and get connected directly with the hiring team

Apply for this Position