Global Production Systems Engineer

Meta

Oregon

On-site

USD 144,000 - 204,000

Full time

11 days ago
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Bonus
Stock options
Benefits

Job summary

Meta is seeking an experienced Production Systems Engineer to join the Data Center Operations team. The role focuses on building automation to improve server uptime and efficiency across a hyperscale fleet.

You will write tooling in Python/Bash/C/C++, analyze telemetry data, and collaborate with global teams to drive operational improvements. Travel up to 25% is required; remote work is not indicated.

Qualifications

  • 6+ years of experience in production systems engineering and large-scale hardware environments.
  • 6+ years of experience with hardware lifecycle management, fleet automation, or data center operations systems.
  • Experience developing systems software or automation tooling in Python, Bash, PHP, C, or C++ for Linux-based production environments at scale.
  • Experience with configuration and maintenance of production systems including web servers, load balancers, relational databases, storage systems, and messaging systems.
  • Experience communicating technical designs and infrastructure decisions through written documentation and cross-functional stakeholder alignment.

Responsibilities

  • Identify and root cause systemic issues across the server fleet to maximize uptime using telemetry.
  • Write, review, and maintain code for diagnostic and automation tooling supporting production servers at hyperscale.
  • Own and develop diagnostic tooling requirements for frontline operations teams.
  • Drive escalation for tooling and hardware issues affecting fleet health.
  • Execute validation activities for new product integration into the production environment.
  • Collaborate with tooling teams to influence roadmaps with an operations-centric view.
  • Perform data analysis to prioritize automation opportunities for server repair workflows.
  • Build cross-functional relationships to improve data center operations efficiency.
  • Mentor engineers on resolving fleet issues and improving tools and processes.
  • Travel up to 25% to support global data center operations and new site deployments.

Skills

Production systems
Fleet automation
Data center operations
Python
Bash
C/C++
Linux production environments
Documentation & collaboration
Data analysis

Education

Bachelor's degree in CS/CE or related
Equivalent practical experience

Tools

Python tooling
Configuration management
Monitoring pipelines

Job description

Summary:

Meta is seeking an experienced Production Systems Engineer to join the Data Center Operations team. Our data centers and the tens of thousands of servers installed within them form the foundation upon which Meta's rapidly scaling infrastructure operates and upon which innovative services are delivered. Meta is at the leading edge of the global data center industry in both design and operations. This role requires a forward-thinking systems professional with deep experience leveraging diverse software tools to identify automation solutions for complex operational challenges. The ideal candidate performs deep data analysis to prioritize server repair automation in a hyperscale environment, drives solutions through code, and collaborates effectively with globally distributed teams through clear written communication.

Required Skills:

Global Production Systems Engineer Responsibilities:

  1. Identify and root cause systemic issues across the server fleet and drive resolutions to maximize uptime and utilization by leveraging hardware failure data and diagnostic telemetry

  2. Write, review, and maintain code for diagnostic and automation tooling that supports quality and efficient delivery of production servers at hyperscale

  3. Own and develop diagnostic tooling requirements that enable frontline operations teams to efficiently manage and repair the server fleet

  4. Drive the escalation process for Data Center Operations to identify, root cause, and resolve complex tooling and hardware issues affecting fleet health

  5. Execute operational validation and verification activities for new product integration into the production environment

  6. Collaborate with cross-functional tooling teams to provide an operations-centric perspective on open issues and contribute to their development roadmaps

  7. Perform deep data analysis to prioritize automation opportunities for server repair workflows in a large-scale, heterogeneous hardware environment

  8. Build cross-functional relationships and influence policies and procedures to improve global data center operations consistency and efficiency

  9. Mentor other engineers on evaluating and resolving fleet issues and defining improvements to tools and operational processes

  10. Travel up to 25% to support global data center operations and new site deployments

Minimum Qualifications:

Minimum Qualifications:

  1. Bachelor's degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience

  2. 6+ years of experience in production systems engineering, infrastructure engineering, or systems software development for large-scale hardware environments

  3. 6+ years of experience with hardware lifecycle management, fleet automation, or data center operations systems spanning compute, storage, or networking infrastructure

  4. Experience developing systems software or automation tooling in Python, Bash, PHP, C, or C++ for Linux-based production environments at scale

  5. Experience with configuration and maintenance of production systems including web servers, load balancers, relational databases, storage systems, and messaging systems

  6. Experience communicating technical designs and infrastructure decisions through written documentation and cross-functional stakeholder alignment across engineering and operations teams

Preferred Qualifications:

Preferred Qualifications:

  1. Experience with data analysis and visualization tools used to prioritize fleet health initiatives and drive operational decision-making

  2. Experience designing or operating configuration management and infrastructure-as-code systems for large heterogeneous hardware fleets

  3. Experience supporting global, multi-site data center infrastructure deployments including hardware qualification and regional rollout coordination

  4. Familiarity with distributed systems monitoring, alerting, and automated remediation pipelines at hyperscale

Public Compensation:

$144,000/year to $204,000/year + bonus + equity + benefits

Industry: Internet

Equal Opportunity:

Meta is proud to be an Equal Employment Opportunity and

Meta is proud to be an Equal Employment Opportunity and

Meta is committed to providing reasonable accommodations for candidates with disabilities in our recruiting process. If you need any assistance or accommodations due to a disability, please let us know at accommodations-ext@meta.com.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Global Production Systems Engineer
Global Production Systems Engineer

Meta • Fort Worth (TX)

On-site
USD 144,000 - 204,000
Global Production Systems Engineer
Global Production Systems Engineer

Meta • Albuquerque (NM)

On-site
USD 144,000 - 204,000
Bonus
Equity
Benefits
Global Production Systems Engineer
Global Production Systems Engineer

Meta • Nebraska

On-site
USD 144,000 - 204,000
Global Production Systems Engineer
Global Production Systems Engineer

Meta • Papillion (NE)

On-site
USD 144,000 - 204,000
Global Production Systems Engineer
Global Production Systems Engineer

Meta • Forest City (NC)

On-site
USD 144,000 - 204,000
Bonus
Equity
Benefits
Global Production Systems Engineer
Global Production Systems Engineer

Meta • Altoona (IA)

On-site
USD 144,000 - 204,000
Bonus
Equity
Benefits
Global Production Systems Engineer
Global Production Systems Engineer

Meta • Aurora (IL)

On-site
USD 144,000 - 204,000
Equity
Global Production Systems Engineer
Global Production Systems Engineer

Meta • Santa Clara (CA)

On-site
USD 144,000 - 204,000
Global Production Systems Engineer
Global Production Systems Engineer

Meta • DeKalb (IL)

On-site
USD 144,000 - 204,000
Bonus potential
Equity
Benefits
Global Production Systems Engineer
Global Production Systems Engineer

Meta • New Albany (OH)

On-site
USD 144,000 - 204,000