OpenAI
OpenAI

Software Engineer, GPU Infrastructure- ChatGPT Engineering

Full-timeLondon, UKApplied AIApplied AI Engineering2mo ago

Job Overview

About the Team

ChatGPT Engineering builds and operates the compute platform powering one of the world's largest AI products. Every ChatGPT conversation relies on massive GPU clusters serving inference workloads with high reliability, efficiency, and performance.

As our GPU fleet continues to grow, we're investing in the infrastructure that operates it. Our team builds the tooling, automation, and intelligent systems that make GPU infrastructure scalable, observable, and increasingly autonomous. We work across production engineering, distributed systems, capacity management, and AI-powered operational tooling to help researchers and product teams move faster while maximizing the efficiency of every GPU.

This is a unique opportunity to work on infrastructure at the frontier of AI, where small improvements in fleet efficiency, reliability, and automation have an outsized impact on the development and deployment of AGI.

About the Role

We're looking for a Software Engineer with deep experience operating large-scale GPU or compute infrastructure.

You'll design and build the systems that manage GPU clusters at scale—from fleet health and capacity planning to operational auto

Core Requirements

Applied AIApplied AI Engineering