ChatGPT Performance Engineer

OpenAI

New York (NY)

On-site

USD 325,000 - 405,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A leading AI research organization in New York is seeking an experienced Performance Engineer to enhance performance, reliability, and efficiency across mission-critical products. The ideal candidate has over 7 years of software engineering experience, particularly in optimizing high-scale distributed systems. Responsibilities include analyzing application performance, developing observability tools, and collaborating with multiple teams to meet performance goals. This role offers the opportunity to impact core infrastructure and improve system metrics significantly.

Qualifications

  • 7+ years of experience in software engineering focused on performance and reliability.
  • Strong understanding of OS internals, scheduling, memory management, and I/O patterns.
  • Comfortable using performance profiling tools and tracing systems.

Responsibilities

  • Analyze and optimize performance across application, middleware, and infrastructure layers.
  • Develop tooling and metrics for system observability.
  • Collaborate with teams to drive systemic improvements in performance.

Skills

Performance profiling tools
High-scale distributed systems
Database optimization
Network optimization
Storage performance
Python/Golang internals

Job description

About the Team

We bring OpenAI's technology to the world through products like ChatGPT and the OpenAI API.

We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. Safety is more important to us than unfettered growth.

About the Role

OpenAI is looking for an experienced Performance Engineer to help us scale the performance, reliability, and efficiency of our systems. In this role, you'll apply deep technical expertise to optimize infrastructure and application-level performance across mission‑critical products like ChatGPT and our developer API. You'll work cross‑functionally with teams building core services, training models, and developing real‑time user experiences to push our latency, throughput, and cost‑efficiency to the next level.

We are looking for engineers who thrive in ambiguous environments, value deep systems understanding, and are motivated by delivering measurable impact. This is a highly technical, individual contributor role focused on root‑cause analysis, profiling, instrumentation, and architecture‑level performance improvements across our stack.

In this role, you will:
  • Analyze and optimize performance across application, middleware, runtime, and infrastructure layers—including networking, storage, Python runtime, GPU utilization, and beyond.
  • Develop tooling and metrics that provide deep observability into system performance.
  • Collaborate closely with infra, platform, training, and product teams to identify key performance goals and drive systemic improvements.
  • Influence architecture and design decisions to prioritize latency, throughput, and efficiency at scale.
  • Lead investigations into high‑impact performance regressions or scalability issues in production.
  • Drive performance testing strategies and help define SLAs/SLOs around latency and throughput for critical systems.
You might thrive in this role if you:
  • Have 7+ years of experience in software engineering with a strong track record in performance or reliability of high‑scale distributed systems.
  • Are deeply comfortable with performance profiling tools and tracing systems.
  • Have experience optimizing performance across one or more layers of the stack (e.g., database, networking, storage, application runtime, GC tuning, Python/Golang internals, GPU utilization).
  • Have a strong understanding of OS internals, scheduling, memory management, and I/O patterns.
  • Have contributed to observability, benchmarking, or performance‑focused infrastructure at scale.
  • Have demonstrated success navigating ambiguity and aligning stakeholders around performance goals.
  • Value simplicity, rigor, and collaboration when solving complex systems problems.

We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic.

Compensation Range: $325K – $405K

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ChatGPT Performance Engineer
ChatGPT Performance Engineer

OpenAI • United States

On-site
USD 325,000 - 405,000
ChatGPT Performance Engineer
ChatGPT Performance Engineer

Slope • San Francisco (CA)

On-site
USD 150,000 - 200,000
ChatGPT Performance Engineer
ChatGPT Performance Engineer

OpenAI • Los Angeles (CA)

On-site
USD 130,000 - 180,000
Software Engineer, ChatGPT Infrastructure
Software Engineer, ChatGPT Infrastructure

OpenAI • San Francisco (CA)

On-site
USD 255,000 - 405,000
Software Engineer, ChatGPT Infrastructure
Software Engineer, ChatGPT Infrastructure

OpenAI • Los Angeles (CA)

Hybrid
USD 255,000 - 405,000
Training Performance Engineer
Training Performance Engineer

Slope • San Francisco (CA)

On-site
USD 250,000 - 460,000
Relocation assistance
Flexible working hours
Collaborative work environment
Frontend Engineer, ChatGPT Engineering
Frontend Engineer, ChatGPT Engineering

OpenAI • Town of Tonawanda (NY)

Hybrid
USD 110,000 - 160,000
Software Engineer, Developer Productivity
Software Engineer, Developer Productivity

OpenAI • San Francisco (CA)

On-site
USD 210,000 - 490,000
Systems Generalist, GPT Infrastructure
Systems Generalist, GPT Infrastructure

Slope • San Francisco (CA)

On-site
USD 180,000 - 260,000
Systems Generalist, GPT Infrastructure
Systems Generalist, GPT Infrastructure

OpenAI • United States

On-site
USD 180,000 - 260,000