Stand out for this role — generate a tailored resume and cover letter in about a minute.
BGC Group Pte Ltd is seeking an experienced IT Infrastructure Administrator to ensure the Bus ETA application infrastructure remains available, performant, and reliable in our Singapore data centre. You will manage physical and virtual servers, monitor health and capacity, and lead incident response, DR testing, and maintenance across OS, virtualization, storage, and hardware.
Candidates should have Red Hat server administration, Podman containers, PostgreSQL, VMware/Hyper-V, storage and backups
Ensure the continuous availability, performance, and reliability of the Bus ETA application infrastructure hosted in the data centre.
Manage and maintain physical and virtual servers supporting the Bus ETA system.
Monitor system health, capacity, and performance to ensure high availability.
Manage, review and ensure the administration work done on the operating systems, virtualization platforms, storage, and server hardware are performed correctly and regularly on time.
Manage, review and ensure the maintenance work done on the UPS systems, server racks, and environmental monitoring within the data centre are performed correctly and regularly on time.
Manage, review and ensure the routine maintenance, patching, firmware upgrades, and lifecycle management are performed correctly and regularly on time.
Manage, review and verify change/service requests as well as oversee that works (during engineering hours) are performed correctly and on time.
Lead incident response and recovery activities for server and infrastructure-related issues.
Develop disaster recovery procedures and conduct failover and restoration testing.
Coordinate with application vendors and stakeholders during planned maintenance and incident resolution.
Produce operational reports, documentation and standard operating procedures.
Experience with administration of Redhat server, Container (such as Podman), PostgreSQL.
Knowledge of virtualization technologies (e.g., VMware or Hyper-V).
Familiarity with storage, backup solutions, and data centre operations.
Experience supporting mission-critical or transport-related systems.
Strong troubleshooting and root cause analysis skills.
Participation in 24/7 standby or on-call support for critical incidents is required.