RefolkCandidates

Software Engineer, ChatGPT Infrastructure

OpenAI · San Francisco

$255k - $405k/yrRemoteMid levelEngineeringApplied AIApplied AI EngineeringPosted 5 months ago

About this role

About the Team ChatGPT is a rapidly evolving system: new capabilities ship continuously, product surfaces change quickly, and usage patterns shift week-to-week. Supporting that pace requires infrastructure that can handle real production constraints - high concurrency, unpredictable traffic patterns, complex dependency graphs, and frequent change. The ChatGPT Infrastructure team builds and operates the platforms that enable fast iteration without compromising performance or reliability. We design shared systems, data paths, rollout mechanisms, and reliability guardrails that teams rely on to ship changes to ChatGPT at scale. We focus on high-leverage infrastructure: primitives and “golden paths” that incorporate operational lessons as defaults, so engineers don’t need to rediscover failure modes, latency pitfalls, or integration issues each time they build something new. About the Role We’re hiring Senior and Staff Engineers to design and build infrastructure systems that underlie ChatGPT and multiply the effectiveness of teams building user experiences. This is not a support-only role. It’s a platform-building role: you’ll define interfaces, develop core abstractions, and create tooling to make safe, fast iteration the norm. Your work will reduce friction, prevent regressions, improve performance, and ensure systems scale gracefully as the product grows. Where You Can Have Impact - You might work on one or more of the following areas (without being restricted to any single area): - Platform foundations & frameworks: Core libraries, service frameworks, and shared components that standardize system building, integration, and evolution. - Scalability & performance primitives: Patterns and infrastructure that reduce tail latency, improve throughput, and keep costs predictable as demand increases. - Reliability guardrails: Mechanisms that prevent outages by design - rate limiting, load shedding, dependency isolation, backpressure, safe fallbacks, and robust regression c

Excerpt from the posting OpenAI published. Read the full description on their site before applying.

Skills this posting mentions

InfrastructureFailure AnalysisDistributed Systems

About OpenAI

OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. AI is an extremely powerful tool that must be created with safety and human needs at its core. OpenAI is dedicated to putting that alignment of interests first - ahead of profit. To achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. Our investment in diversity, equity, and inclusion is ongoing, executed through a wide range of initiatives, and championed and supported by leadership. At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.

All 755 openings at OpenAI

Apply to this role, tailored

Queue Software Engineer, ChatGPT Infrastructure at OpenAI and I will read the posting, rewrite your resume against it, draft the cover letter, and score the fit before you send anything.

Applications finish as drafts. Nothing is sent until you read it and press send. New accounts start with 500 free credits.

More roles at OpenAI

See all

Similar roles elsewhere

See more

Listed from the job board OpenAI publishes. Refolk is not the employer and does not handle their hiring. Applications go to OpenAI directly.