Fleet Operations Manager, Data Center Infrastructure

Meta Careers

Montgomery (AL)

On-site

USD 163,000 - 238,000

Full time

11 days ago
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Bonus
Equity
Benefits

Job summary

Meta is seeking a forward-thinking Fleet Operations Manager to lead a geographically dispersed data center team. You will deliver SLA/KPI targets for production server hardware across a large-scale fleet in a 24/7 environment.

You will develop engineers, drive continuous improvement, and act as incident manager during major events while leveraging data analytics to identify issues and opportunities for optimization.

Qualifications

  • BS/BA/BEng in a technical field or commensurate experience.
  • Experience leading technical projects related to process improvement, technology, and automation.
  • 5+ years of experience managing teams of technical resources and performance.
  • Understanding of data center infrastructure including power, cooling, and networks.

Responsibilities

  • Build and lead a geographically dispersed data center operations team.
  • Serve as incident manager during large-scale production events and coordinate cross-functional response.
  • Analyze data to identify inefficiencies and drive continuous improvement.
  • Collaborate with partner teams to maintain fleet health and capacity levels.

Skills

Server hardware
Project management
Quality management
Data analytics
Networks
OS repair
Linux and automation

Education

BS/BA/BEng in technical field

Job description

Meta is seeking a forward-thinking, experienced individual to join the Data Center Fleet Operations team. The Fleet Operations Manager is accountable for managing and leading a geographically dispersed team, delivering SLA/KPI’s related to production server hardware, resolution of systemic technical issues, and repairs throughout the assigned geographic region of data centers. We are looking for someone who can prioritize competing work based on impact, deadlines, and stakeholder needs in a 24/7 data center operations environment. The ideal candidate is an IT professional with strong leadership skills and experience in Server Hardware, Project Management, Quality Management, Data Analytics, Networks, OS repair, Linux and Automation, ideally in a datacenter environment. Experience managing servers in a large-scale distributed environment is important for success in this roleFleet Operations Manager, Data Center Infrastructure Responsibilities:Build and lead a geographically dispersed, high-performing data center operations team, developing both the technical capabilities and leadership qualities of engineersEstablish and manage a Data Center Operations Team accountable for the maintenance and operation of server hardware and supporting infrastructure at scaleBecome a technical expert in Meta's infrastructure, including platforms, tools, systems, architecture, workflows, and performanceProvide strategic direction, guidance, and support for site and fleet-level operationsAnalyze and drive continuous improvement in the engineering and operational performance of our data centersEmploy data analytics to identify inefficiencies, opportunities, exceptions, and correlations in a complex, highly interconnected, technical environment. Enable rapid and effective problem solving, along with proactive identification and mitigation of risks and issuesCollaborate with cross-functional partner teams to ensure fleet health and maintain targeted capacity levels, resulting in optimized operations, minimized downtime, and seamless scalabilityEvolve and optimize processes in a globally consistent way to allow Meta to scale and grow effectivelySupport and mentor engineers in their day-to-day work, as well as in finding opportunities to develop and grow based on their areas of strength and interestFoster an environment of ownership, innovation, collaboration, accountability, continuous improvement, and safetyConduct performance management for a technical engineering team, providing clear expectations and goalsAssume the role of incident manager during large-scale, site-wide, and region-wide production-impacting events, as the primary point of contact for your site. This requires working cross-functionally to scope problems, mitigate risks, affect fixes, and communicate the nature, status, and resolution plan for incidentsSupport and contribute thought leadership to the development and implementation of business practices, processes and automated toolingDevelop deep knowledge and ownership of a hyper-scale computing fleet through the use of data analysis to identify trends and systemic issues and opportunitiesreporting out globally and sharing with peers as appropriateMinimum Qualifications:BS, BA, or BEng in a technical field or commensurate experienceAbility to travel up to 30% is requiredExperience participating in or leading technical projects related to areas such as process improvement, technology, and/or automation, including bringing in additional expertise as needed5+ years of experience managing teams of technical resources, including people and performance management responsibilitiesUnderstanding of data center infrastructure and/or operations, including power, cooling, and/or network systemsstructured cablingand management of projects, incidents, and vendorsExperience using data and metrics to drive decision-makingAbility to influence effectively, working on cross-functional teams to advance the needs of the company and adapting teams to meet these needs10+ years of engineering or operations experience, preferably in a mature engineering or operations environment, working with cross-functional teamsAbility to communicate effectively, in a clear and concise manner, appropriately tailoring messages to the audiencePreferred Qualifications:Experience leading technical resources using Linux or an equivalent OS to support hardware systems in a complex IT environmentExperience in large-scale data center hardware deployments and building scalable infrastructureExperience adhering to and implementing responsible, ethical AI practices (e.g., risk assessment, bias mitigation, quality and accuracy reviews)Knowledge of the interdependencies of data center functions and technologies, including electrical, cooling, structured cabling, security, and networkSix Sigma knowledge/certificationDemonstrated ability to integrate AI tools to optimize/redesign workflows and drive measurable impact (e.g., efficiency gains, quality improvements)Experience with large-scale AI implementations and the use of AI to drive automationDemonstrated ongoing AI skill development (e.g., prompt/context engineering, agent orchestration) and staying current with emerging AI technologiesAbout Meta:Meta builds technologies that help people connect, find communities, and grow businesses. When Facebook launched in 2004, it changed the way people connect. Apps like Messenger, Instagram and WhatsApp further empowered billions around the world. Now, Meta is moving beyond 2D screens toward immersive experiences like augmented and virtual reality to help build the next evolution in social technology. People who choose to build their careers by building with us at Meta help shape a future that will take us beyond what digital connection makes possible today—beyond the constraints of screens, the limits of distance, and even the rules of physics.Meta is proud to be an Equal Employment Opportunity and Affirmative Action employer. We do not discriminate based upon race, religion, color, national origin, sex (including pregnancy, childbirth, or related medical conditions), sexual orientation, gender, gender identity, gender expression, transgender status, sexual stereotypes, age, status as a protected veteran, status as an individual with a disability, or other applicable legally protected characteristics. We also consider qualified applicants with criminal histories, consistent with applicable federal, state and local law. Meta participates in the E-Verify program in certain locations, as required by law. Please note that Meta may leverage artificial intelligence and machine learning technologies in connection with applications for employment.Meta is committed to providing reasonable accommodations for candidates with disabilities in our recruiting process. If you need any assistance or accommodations due to a disability, please let us know at accommodations-ext@meta.com.$163,000/year to $238,000/year + bonus + equity + benefitsIndividual compensation is determined by skills, qualifications, experience, and location. Compensation details listed in this posting reflect the base hourly rate, monthly rate, or annual salary only, and do not include bonus, equity or sales incentives, if applicable. In addition to base compensation, Meta offers benefits. Learn more about benefits at Meta.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Data Center Lead - Houston
Data Center Lead - Houston

Meta Careers • Houston (TX)

On-site
USD 111,000 - 159,000
Bonus
Equity
Benefits
Production Systems Engineer, Fleet AI Systems
Production Systems Engineer, Fleet AI Systems

Meta • Menlo Park (CA)

On-site
USD 173,000 - 245,000
Bonus
Equity
Benefits
Data Center Capacity Engineer
Data Center Capacity Engineer

Meta • Forest City (IA)

On-site
USD 84,000 - 130,000
Bonus
Equity
Benefits
Global Production Platform Engineer
Global Production Platform Engineer

Meta Careers • Menlo Park (CA)

On-site
USD 144,000 - 204,000
Data Center Production Operations Engineer
Data Center Production Operations Engineer

Meta • Oregon (WI)

On-site
USD 71,000 - 103,000
Bonus
Equity
Benefits
Finance Manager, Infrastructure Data Center
Finance Manager, Infrastructure Data Center

Meta • Menlo Park (CA)

On-site
USD 120,000 - 175,000
Bonus
Equity
Benefits
Data Center Infrastructure Management (DCIM) Engineer
Data Center Infrastructure Management (DCIM) Engineer

Meta Careers • Reston (VA)

On-site
USD 137,000 - 200,000
Capacity Engineer, Infrastructure Data Governance
Capacity Engineer, Infrastructure Data Governance

Meta Careers • Menlo Park (CA)

On-site
USD 154,000 - 217,000
Bonus
Equity
Benefits
Data Center Capacity Manager
Data Center Capacity Manager

Meta • DeKalb (IL)

On-site
USD 160,000 - 223,000
Bonus
Equity
Benefits
Global Production Systems Engineer
Global Production Systems Engineer

Meta • Georgia

On-site
USD 144,000 - 204,000
Bonus
Equity
Benefits