DCEO Principal Engineer, Infrastructure Operations

Amazon Web Services (AWS)

Mumbai

On-site

INR 3,000,000 - 6,000,000

Full time

33 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

ADSIPL in Maharashtra is seeking a Principal Engineer to own site operations for data-center facilities, coordinating with cross-functional teams to ensure safety, security and availability. You will lead the engineering operations, drive improvements, and represent India in global initiatives.

The role emphasizes managing complex electrical and mechanical systems, incident response, and Kaizen-driven performance enhancements, with a focus on building scalable processes and leadership

Qualifications

  • 12+ years of management experience in data-center or mission-critical facilities.
  • Master's degree in Electrical Engineering, Mechanical Engineering, or related field.
  • Knowledge of electrical and mechanical systems in critical data center operations.
  • Experience hiring, developing, and managing technical teams.
  • Experience with large-scale mechanical and power systems.

Responsibilities

  • Own as the site SME and POC for mission-critical facilities including vendor management and site operations.
  • Plan, review, and drive capacity, availability, and other projects across assigned sites.
  • Prepare countermeasures for natural disasters and respond to high-severity incidents and emergencies.
  • Build team structures, recruit, develop, and mentor builders for long-term excellence.

Skills

Management experience
Team leadership
Power systems

Education

Master's degree in Electrical/Mechanical Engineering

Job description

DESCRIPTION
DESCRIPTION

AWS Infrastructure Services owns the design, planning, delivery, and operation of all AWS global infrastructure. In other words, we’re the people who keep the cloud running. We support all AWS data centers and all of the servers, storage, networking, power, and cooling equipment that ensure our customers have continual access to the innovation they rely on. We work on the most challenging problems, with thousands of variables impacting the supply chain — and we’re looking for talented people who want to help.

You’ll join a diverse team of software, hardware, and network engineers, supply chain specialists, security experts, operations managers, and other vital roles. You’ll collaborate with people across AWS to help us deliver the highest standards for safety and security while providing seemingly infinite capacity at the lowest possible cost for our customers. And you’ll experience an inclusive culture that welcomes bold ideas and empowers you to own them to completion.

The Amazon Data Centre Engineering Operation ( DCEO ) team is seeking a strong subject matter expert in electrical/mechanical/control and firefighting system as a Principal Engineer in India Infra operation team. Responsibilities of the Principal Engineer for India region are to deploy solutions for complex build and operations related issues for the DCEO domain. Job description contains a comprehensive job requirement for the Principal Engineer, Data Centre Engineering Operations. The additional responsibilities are: 1/Support the process definition and policies, supply procedures for emergency and other standard requirements, and communicate the contexts in which the rules, processes, and polices are applied. 2/Inspect and validate the requirements and deliverables that describe the build and service (s). 3/Give input to design and construction during the building phase, as well as feedback from operations. 4/Provide input into and execute user documentation and training material. 5/Test the new build (s) or service (s) under the process, including post-test script review (user acceptance testing), by using and evaluating it for accuracy and usability. 6/Get or provide approval for deviations or changes to rules, processes, and policies. 7/Function as the escalation point for all operations-related issues. 8/Review Ops metrics and establish procedures and a path to green for keeping them on track. 9/Help troubleshoot facility and rack-level events within internal Service Level Agreements (SLA). 10/Help the regular operations team with technical review and support in COE preparation in faulty scenarios. 11/Support in the rectification of major technical changes by liaising with Design & Field Engineering. 12/Support new product implementation by attending game days and providing training cluster wide. 13/Assist in projects for continuous improvement and operational excellence. 14/Own cluster-wide availability goals like MOT, LLT, and LSE drill compliance. 15/Represent India in global level forums.

Main Responsibilities

Own as the site SME and POC, plan, review, evaluate, operate, maintain, improve and manage mission-critical facilities including vendor management, day to day hands-on work and supervision relating to decrease/increase of rack capacity, onsite on-going or future construction works, planned maintenance works and urgent or emergency changes along with the AWS Infrastructure Priorities.

Participate in and be responsible for future Capacity, Availability and other projects of assigned sites, review, evaluate and give feedback on designs from Operations view point to mitigate Safety, Security and Availability risks beforehand.

Prepare and implement countermeasure for natural disasters, emergency response to high priority/critical incidents including creating EOPs, training staffs and preparing appropriate tools. Respond to high severity events and large scale event as the owner of the operations. Understand SOO and EOPs, troubleshoot, mitigate, and resolve issues, write and update senior leaders through regular and timely report, conclude issue with complete root cause analysis.

Review, evaluate and proactively identify SPOF risks or vulnerability in data centre (electrical, mechanical, control) designs, test and commissioning program, construction and operations processes, and consider, plan, coordinate, propose, negotiate, persuade, grant approval from stakeholders for the issue remediation and/or mitigation plan and deliver results.

Build sustainable and scalable mechanism to collect, review and report regular metrics and KPI of the team, plan, propose, and drive kaizen based on the metrics and KPI results.

Understand and develop team structure, create and document headcount requirements, help drive interview, and hire bar raising candidates, build strong team through delegation, development, training, directing, coaching, empowering and motivating builders.

Basic Qualifications

10 years+ experience with designing, building, commissioning, operating and maintaining data centre or mission critical facilities such as power substation, airport, hospital, etc.

Experience in managing life cycle of data centre from designing, constructing, commissioning, operating and to decommissioning.

Has strong ability to understand electrical systems (be able to read and write SLD with no information in hand (Single Line Diagram with all details including breaker open / close state and SOO)), (supply system of power substations, transformers, switchgears, VFI-class UPS, DRUPS, PDU, ATS, STS, SLA or VRLA battery and related systems, fuel systems related to diesel/gas turbine generators, surge control circuits, active harmonic filters, battery monitoring systems, branch circuit monitoring systems, SCADA systems)

Has strong ability to understand mechanical systems (CRAC/CRAH, AHU, chillers, cooling towers, storage tanks, heat exchangers, plumbing systems, pumps, valves, duct systems, fans, dampers, fire detection and extinguishing systems, drainage systems, building monitoring systems, automatic control systems).

Experience and deep understanding with change management, incident management, problem management (including troubleshooting incident as incident commander and post failure/root cause analysis), vendor management, risk management, asset management (critical equipment and spares), energy management (PUE improvement, government reporting), BCP (business continuity planning), annual and mid to long-term maintenance planning and management, budgeting, reporting, communicating with senior leaders, PDCA cycle for process improvement, etc.

Experience in managing and operating multiple data centre sites.

Be able to represent the country for the role for global initiatives, work with cross functional teams and deliver expected results in a timely manner.

Key job responsibilities

Own as the site SME and POC, plan, review, evaluate, operate, maintain, improve and manage mission-critical facilities including vendor management, day to day hands-on work and supervision relating to decrease/increase of rack capacity, onsite on-going or future construction works, planned maintenance works and urgent or emergency changes along with the AWS Infrastructure Priorities.

Participate in and be responsible for future Capacity, Availability and other projects of assigned sites, review, evaluate, and give feedback on designs from Operations view point to mitigate Safety, Security and Availability risks beforehand.

Prepare and implement countermeasure for natural disasters, emergency response to high priority/critical incidents including creating EOPs, training staffs and preparing appropriate tools. Respond to high severity events and large scale event as the owner of the operations. Understand SOO and EOPs, troubleshoot, mitigate, and resolve issues, write and update senior leaders through regular and timely report, conclude issue with complete root cause analysis.

Review, evaluate and proactively identify SPOF risks or vulnerability in data centre (electrical, mechanical, control) designs, test and commissioning program, construction and operations processes, and consider, plan, coordinate, propose, negotiate, persuade, grant approval from stakeholders for the issue remediation and/or mitigation plan and deliver results.

Build sustainable and scalable mechanism to collect, review and report regular metrics and KPI of the team, plan, propose, and drive kaizen based on the metrics and KPI results.

Understand and develop team structure, create and document headcount requirements, help drive interview, and hire bar raising candidates, build strong team through delegation, development, training, directing, coaching, empowering and motivating builders.

About The Team
About AWS
Diverse Experiences

Amazon values diverse experiences. Even if you do not meet all of the preferred qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn’t followed a traditional path, or includes alternative experiences, don’t let it stop you from applying.

Why AWS

Amazon Web Services (AWS) is the world’s most comprehensive and broadly adopted cloud platform. We pioneered cloud computing and never stopped innovating — that’s why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses.

Work/Life Balance

We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why we strive for flexibility as part of our working culture. When we feel supported in the workplace and at home, there’s nothing we can’t achieve.

Inclusive Team Culture

AWS values curiosity and connection. Our employee-led and company-sponsored affinity groups promote inclusion and empower our people to take pride in what makes us unique. Our inclusion events foster stronger, more collaborative teams. Our continual innovation is fueled by the bold ideas, fresh perspectives, and passionate voices our teams bring to everything we do.

Mentorship and Career Growth

We’re continuously raising our performance bar as we strive to become Earth’s Best Employer. That’s why you’ll find endless knowledge-sharing, mentorship and other career-advancing resources here to help you develop into a better-rounded professional.

BASIC QUALIFICATIONS
  • 12+ years of management experience
  • Master's degree in Electrical Engineering, Mechanical Engineering, or a related field
  • Knowledge of the electrical and mechanical systems involved in critical data center operations including systems such as feeders, transformers, generators, switchgear, UPS systems, ATS units, PDU units, chillers, pumps, air handling units, and CRAC units
  • Experience with large scale mechanical and power systems
  • Experience hiring, developing, and managing high-performing technical teams
PREFERRED QUALIFICATIONS
  • Knowledge of electrical engineering best practices including breaker coordination studies, switchgear sequence of operation, and NEC code
  • Experience owning the operation of a mission-critical team or product
  • Experience with large-scale technical operations or large-scale compute farms
  • Experience with process improvement techniques such as Kaizen, Lean Manufacturing or Six Sigma

Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how-we-hire/accommodations for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.

Company - ADSIPL - Maharashtra

Job ID: A10515197

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

DCEO Principal Engineer, Infrastructure Operations
DCEO Principal Engineer, Infrastructure Operations

Amazon • Mumbai

On-site
INR 4,000,000 - 7,000,000
Data Centre Engineering Operations Engineer, Data Centre Engineering Operations
Data Centre Engineering Operations Engineer, Data Centre Engineering Operations

Amazon Web Services (AWS) • Maharashtra

On-site
INR 1,500,000 - 2,500,000
Data Centre Engineering Operations, Data Centre Engineering Operations
Data Centre Engineering Operations, Data Centre Engineering Operations

Amazon Web Services (AWS) • Maharashtra

On-site
INR 1,500,000 - 2,500,000
Data Center Engineering Operations , DCEO
Data Center Engineering Operations , DCEO

Amazon Web Services (AWS) • Mumbai

On-site
INR 900,000 - 1,300,000
Data Centre Engineering Operations Cluster Manager, AWS Data Center Operations
Data Centre Engineering Operations Cluster Manager, AWS Data Center Operations

Amazon Web Services (AWS) • Mumbai

On-site
INR 4,000,000 - 6,000,000
Facility Operations Center Engineer
Facility Operations Center Engineer

Amazon Web Services (AWS) • Mumbai

On-site
INR 600,000 - 1,200,000
Data Center Engineering Operations Engineer, DCEO, BOM
Data Center Engineering Operations Engineer, DCEO, BOM

Amazon Web Services (AWS) • Mumbai

On-site
INR 900,000 - 1,400,000
Facility Operations Center Engineer, DCC Communities
Facility Operations Center Engineer, DCC Communities

Amazon Web Services (AWS) • Maharashtra

On-site
INR 900,000 - 1,500,000
Data Centre Engineering Operations, Data Centre Engineering Operations
Data Centre Engineering Operations, Data Centre Engineering Operations

Amazon • Mumbai

On-site
INR 1,500,000 - 2,100,000
DCEO Engineer
DCEO Engineer

Amazon • Mumbai

On-site
INR 2,500,000 - 4,500,000