Princeton Digital Group (PDG) is a leading developer and operator of AI and cloud hyperscale data centers in Asia Pacific. Headquartered in Singapore, PDG has a presence across high-growth markets in the region including Singapore, Japan, India, Indonesia, China, Malaysia, South Korea, and Australia
PDG is backed by some of the world’s most reputed blue-chip investors - Warburg Pincus, a leading global private equity firm with a successful history of creating world-class technology, media, and telecommunications (TMT) and real estate platforms; Ontario Teachers' Pension Plan (OTPP), Canada's largest single profession pension plan; Mubadala Investment Company, a sovereign investor that manages a diverse global portfolio for the Government of Abu Dhabi and Stonepeak, a leading alternative investment firm specializing in infrastructure and real assets globally.
Why Join PDG?
- Impact at Scale: Be part of a company that’s enabling digital transformation of economies.
- Career Growth Without Borders: With operations across Asia Pacific and a fast-growing footprint, your next opportunity could be anywhere.
- Culture That Empowers: We foster a collaborative, inclusive environment where every voice is heard and every contribution matters.
- Commitment to Excellence: From governance to workplace safety, we hold ourselves to the highest standards because our people deserve nothing less.
PDG empowers individuals in every role to grow, lead, and leave a lasting impact.
Join us to build the digital future of Asia.
For more information, visit our website www.princetondg.com or follow us on LinkedIn.
We are looking for
As the Senior Critical Systems Manager (BMS/EPMS), you will report to the Head of Operations and lead the management and continual improvement of data center infrastructure systems across new or live environments. You will own key programs spanning capacity and asset visibility, maintenance governance, reliability performance, and risk/compliance to support safe, highly available operations.
To succeed in this role, you should have strong technical depth in mechanical and electrical systems (including fire and security) and the leadership, communication, and organizational skills to operate in a mission-critical environment.
Job Responsibilities
- Own and govern the CMMS program, including asset hierarchy, preventive maintenance (PM) library, work order management, failure coding, and reporting frameworks, to drive PM compliance, reduce repeat failures, and minimize unplanned downtime across critical electrical, mechanical, and fire protection systems.
- Develop and continuously optimize maintenance strategies (preventive, predictive, and reactive) using failure trend analysis, condition monitoring data, and OEM guidelines to improve asset reliability and lifecycle cost efficiency.
- Establish, track, and drive operational KPIs, including availability, MTTR, PM completion, backlog health, and repeat failure rates; lead corrective actions to ensure sustained performance against SLA and uptime targets.
- Lead reliability engineering practices, including root cause analysis (RCA) and corrective/preventive action (CAPA) programs, to eliminate systemic issues and strengthen infrastructure resilience.
- Manage performance of contractors and in-house engineering teams, ensuring high-quality work order execution, SLA adherence, and compliance with site procedures, safety standards, and contractual obligations through regular audits and governance.
- Own the lifecycle management of BMS/CCCS systems, including requirements definition, configuration standards, change control, and validation, ensuring alignment with data center performance, redundancy, and availability requirements.
- Review and approve control logic, including sequences of operation, setpoints, alarms, and interlocks; ensure accurate documentation, version control, and safe implementation of all changes.
- Drive continuous improvement initiatives to enhance system reliability, automation, monitoring visibility, and energy efficiency across critical infrastructure and control environments.
- Plan and execute system upgrades and lifecycle replacements, including firmware updates and enhancements, under formal change management processes with risk assessments, testing, validation, and controlled maintenance windows.
- Ensure OT and BMS/CCCS cybersecurity compliance, including network segmentation, access control, logging, patching, and vulnerability management; support internal and external audits and remediation activities.
- Oversee DCIM platforms and operational dashboards, ensuring accurate, real-time visibility of capacity utilization, asset performance, alarms, and infrastructure health to support operational decision-making.
- Leverage advanced analytics and digital tools (e.g., anomaly detection, predictive maintenance models) to proactively identify risks, improve early fault detection, and reduce unplanned outages.
Requirements
- Diploma or Degree in Electrical Engineering, Mechanical Engineering, Building Services, or a related discipline.
- Relevant experience in data center or mission-critical facilities (24x7 operations environment) with demonstrated ownership of maintenance, reliability, and systems (BMS/CMMS/DCIM):
- 10–15 years for Senior Critical Systems Manager (BMS/EPMS), including broader site or multi-site responsibility.
- Hands-on experience with CMMS and maintenance strategy execution, including asset structuring, PM program development, work order quality, and failure tracking.
- Strong working knowledge of BMS/controls systems, including exposure to:
- Sequences of operations, alarms, and interlocks
- Change management and system validation
- Involvement in commissioning or system upgrades
- Demonstrated experience in reliability engineering practices, including root cause analysis (RCA), failure trend analysis, and implementation of corrective/preventive actions (CAPA).
- Data-driven mindset, with the ability to analyze operational data (e.g., failure trends, PM compliance, MTTR) and translate insights into actionable improvements; experience with tools such as Power BI or equivalent is preferred.
- Familiarity with DCIM, ERP systems, and reporting tools to support asset visibility, capacity tracking, and operational performance monitoring.
- Strong understanding of operational governance, including adherence to MOP/SOP/EOP, change management processes, and audit/compliance requirements in a critical environment.
- Proven ability to manage vendors and internal teams, ensuring SLA adherence, quality of execution, and compliance with safety and operational standards.
- Ability to communicate technical concepts effectively to both technical and non-technical stakeholders.