Software Engineer, Infrastructure - Core Experimentation

OpenAI

Bellevue (WA)

Hybrid

USD 293,000 - 325,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Relocation assistance

Job summary

OpenAI’s Statsig team seeks a Senior Infrastructure Engineer to design and scale the foundations behind experimentation and rollout. You will work on the distributed control plane, SDK and server evaluation paths, and analytics infrastructure to enable safe, measurable launches at OpenAI scale.

This role requires deep technical excellence in performance, scalability, reliability, and data availability. You will partner with multiple OpenAI product teams to deliver durable platform capabilities

Qualifications

  • Experience building large-scale distributed systems with strict performance, availability, and correctness requirements.
  • Strong background in low-latency configuration delivery, SDK/runtime performance, and high-throughput service design.
  • Experience with data pipelines, analytics systems, and operational tooling for reliable, scalable platforms.

Responsibilities

  • Design and operate low-latency configuration delivery systems powering feature flags and progressive rollouts.
  • Scale SDK, server-side evaluation, and control-plane systems for high-volume services.
  • Build high-throughput data ingestion and analytics infrastructure for experimentation and product analytics.
  • Improve performance, reliability, and observability of core infrastructure as OpenAI products scale.
  • Lead architecture initiatives across ChatGPT, Codex, ads, subscriptions, and developer products.

Skills

Distributed systems
Low-latency
High-throughput
Reliability
Observability
Caching
Concurrency

Education

Bachelor's or Master's in Computer Science

Job description

About the Team

The Statsig team within OpenAI owns the experimentation, rollout, dynamic configuration, and analytics infrastructure that sits on the launch path for OpenAI products. Our systems help teams ship safely, evaluate product and model changes in production, and make high‑confidence decisions from real‑world usage.

Statsig began as an independent company built around experimentation, feature management, and product analytics at scale. After Statsig joined OpenAI, the team began the next chapter: bringing that platform expertise and infrastructure into OpenAI as the experimentation and rollout foundation for every product we ship.

This is infrastructure with a very direct product consequence. Teams working on ChatGPT, Codex, model measurement, consumer experiences including ads, business subscriptions, developer products, and shared platform systems depend on Statsig to evaluate configurations, move traffic safely, ingest experiment data, serve analytics, and roll changes forward or back when production reality demands it.

We are at a critical point in the platform journey. Adoption is accelerating quickly across OpenAI, and the systems that were already important are becoming load‑bearing for how the company launches. The infrastructure needs to stay fast under sharply increasing evaluation volume, reliable when more services depend on it, observable enough to debug quickly, and efficient enough to support OpenAI‑wide scale.

Recent SDK and server‑side infrastructure work has already produced measurable wins in latency, reliability, memory usage, and compute efficiency across important services. The next phase is to make those gains systematic: a platform that can absorb rapidly growing product velocity while preserving low latency, data quality, operational safety, and developer trust.

Based out of OpenAI's Bellevue office, we are a close‑knit team that values in‑person collaboration, technical depth, operational ownership, and building infrastructure that lets other builders move faster without taking on hidden reliability risk.

About the Role

As a Senior Infrastructure Engineer on the Statsig team, you will build and scale the foundational systems behind OpenAI's experimentation and rollout platform. You will work on the distributed control plane, SDK and server evaluation paths, ingestion pipelines, analytics foundations, and operational tooling that make launches safe and measurable at OpenAI scale.

This role is deeply technical and centered on performance, scalability, reliability, and correctness. You will design systems that serve low‑latency configuration decisions, handle high‑throughput event ingestion, preserve data availability for experimentation and analytics workflows, and keep critical launch infrastructure dependable as usage grows.

The work matters because OpenAI's next phase depends on learning quickly without compromising safety or reliability. Every major product surface needs a trusted way to evaluate changes, progressively roll them out, understand impact, and recover cleanly. Statsig is one of the core infrastructure layers that makes that possible.

In This Role, You Will
  • Design and operate low‑latency configuration delivery systems powering feature flags, dynamic configs, and progressive rollouts across OpenAI product suites.
  • Scale SDK, server‑side evaluation, and control‑plane systems so high‑volume services can depend on Statsig without adding user‑visible latency or operational fragility.
  • Build high‑throughput data ingestion and analytics infrastructure for experimentation, product analytics, feature performance monitoring, and model or product measurement workflows.
  • Improve performance, efficiency, reliability, and observability of core Statsig infrastructure as OpenAI products scale globally.
  • Optimize query performance, data freshness, and data availability for teams making launch decisions from experimentation and analytics workflows.
  • Strengthen operational excellence through better SLOs, alerting, debugging tools, incident response, capacity planning, and failure‑mode design.
  • Partner with teams across ChatGPT, Codex, model measurement, consumer ads, business subscriptions, developer products, and infrastructure to turn recurring launch and measurement needs into durable platform capabilities.
  • Lead large technical initiatives and shape the architecture of experimentation and rollout infrastructure used across the company.
You Might Thrive In This Role If You
  • Have experience building large‑scale distributed systems with strict performance, availability, and correctness requirements.
  • Enjoy low‑latency systems work, including real‑time configuration delivery, SDK/runtime performance, caching, concurrency, and high‑throughput service design.
  • Have built or operated large‑scale data platforms, event pipelines, analytics systems, or query infrastructure where freshness and correctness matter.
  • Care deeply about reliability, observability, incident response, capacity planning, and the practical craft of keeping critical production systems boring in the best way.
  • Are excited by the post‑acquisition chapter of a high‑performing infrastructure team: preserving Statsig's strengths while integrating deeply into OpenAI's product and platform stack.
  • Want to work on infrastructure that directly affects how ChatGPT, Codex, model measurement, ads, business subscriptions, developer products, and future OpenAI surfaces ship.
  • Take ownership of complex technical problems end to end and enjoy building infrastructure that lets other teams move faster with more confidence.
Location

This role is based in Bellevue, WA. The team works in person three days per week and offers relocation assistance to new employees.

About OpenAI

OpenAI is an AI research and deployment company dedicated to ensuring that general‑purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity.

We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic.

OpenAI’s affirmative action and equal employment opportunity policy statement is available in the official policy documents.

Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US‑based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non‑public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations.

We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via the provided link.

OpenAI global applicant privacy policy applies.

At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.

Compensation
$293K – $325K + Offers Equity
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Software Engineer, Product - Core Experimentation
Software Engineer, Product - Core Experimentation

OpenAI • Bellevue (WA)

On-site
USD 293,000 - 324,000
Software Engineer, Product - Core Experimentation
Software Engineer, Product - Core Experimentation

OpenAI • Seattle (WA)

Hybrid
USD 293,000 - 324,000
Relocation assistance
Engineering Manager, Core Experimentation
Engineering Manager, Core Experimentation

OpenAI • Seattle (WA)

On-site
USD 250,000 - 320,000
Software Engineer, Product - Core Experimentation
Software Engineer, Product - Core Experimentation

Slope • Seattle (WA)

On-site
USD 170,000 - 250,000
Engineering Manager, Core Experimentation
Engineering Manager, Core Experimentation

OpenAI • United States

On-site
USD 180,000 - 240,000
Software Engineer, ChatGPT Infrastructure
Software Engineer, ChatGPT Infrastructure

OpenAI • San Francisco (CA)

On-site
USD 255,000 - 405,000
Software Engineer, ChatGPT Infrastructure
Software Engineer, ChatGPT Infrastructure

OpenAI • Los Angeles (CA)

Hybrid
USD 255,000 - 405,000
Software Engineer, Build Systems / CI
Software Engineer, Build Systems / CI

OpenAI • New York (NY)

On-site
USD 185,000 - 490,000
Data Engineer, Core Experimentation
Data Engineer, Core Experimentation

OpenAI • Bellevue (WA)

Hybrid
USD 293,000 - 325,000
Data Engineer, Core Experimentation
Data Engineer, Core Experimentation

Slope • Seattle (WA)

Hybrid
USD 293,000 - 325,000