Get more replies from employers
Send a job-specific resume in minutes.
MINDTECK SINGAPORE PTE LTD is seeking an experienced Platform Infrastructure Lead to manage and operate enterprise container platforms and COTS components in a tenant environment above the hypervisor layer, ensuring reliability, security, and operational excellence.
You will lead a team of infrastructure engineers, drive continuous improvements, and collaborate with application and AFE teams to support deployments and issue resolution across mission-critical systems in Singapore.
We are seeking an experienced Platform Infrastructure Lead to manage and operate enterprise platform services within a tenant environment above the hypervisor layer. The successful candidate will be responsible for ensuring the reliability, security, availability, and operational excellence of container platforms and Commercial Off-The-Shelf (COTS) components supporting mission‑critical applications.
This role requires strong technical leadership to manage platform operations, drive continuous improvements, and lead a team of infrastructure engineers while collaborating with application and infrastructure stakeholders.
Lead and manage the Platform Operations team supporting container platforms and COTS services.
Ensure the availability, reliability, security, and performance of platform services.
Manage platform incidents, service requests, changes, problem management, and Root Cause Analysis (RCA).
Install, configure, administer, and troubleshoot COTS components, including MongoDB, Microsoft SQL Server, Nginx, Apache HTTP Server, IBM MQ, MQTT, ArcGIS, and similar enterprise platforms.
Collaborate closely with Application teams and Authority Furnished Equipment (AFE) teams to support system integration, deployments, and issue resolution.
Develop, maintain, and continuously improve Standard Operating Procedures (SOPs), operational runbooks, and platform architecture documentation.
Implement and maintain monitoring, logging, alerting, and observability solutions to ensure proactive platform management.
Plan and execute platform patching, upgrades, vulnerability remediation, and lifecycle management activities.
Ensure operating systems comply with security hardening standards, including golden image compliance and security baseline requirements.
Perform capacity planning, performance tuning, and resource optimization to support business growth.
Lead Disaster Recovery (DR) and Business Continuity Planning (BCP) activities, including regular testing and validation.
Drive operational excellence through automation, process improvements, and adoption of infrastructure best practices.
Mentor and provide technical leadership to a team of infrastructure engineers.
Bachelor's Degree in Computer Science, Information Technology, Computer Engineering, or a related discipline.
7–10 years of experience in infrastructure, platform operations, or systems engineering.
Proven experience managing and supporting enterprise Kubernetes or container platform environments.
Hands‑on experience with container technologies such as Docker and Podman.
Strong experience installing, configuring, administering, and troubleshooting enterprise COTS software, including:
MongoDB
Microsoft SQL Server
Nginx
Apache HTTP Server
IBM MQ
MQTT
ArcGIS
or equivalent enterprise platforms.
Experience with enterprise messaging technologies such as IBM MQ, MQTT, Kafka, or equivalent messaging platforms.
Good understanding of networking concepts, including firewalls, load balancers, DNS, and SSL/TLS, with the ability to troubleshoot platform connectivity issues.
Strong Linux system administration skills, including OS hardening, security configuration, and system performance tuning.
Experience with monitoring and observability platforms such as eG Enterprise, Grafana, Prometheus, ELK Stack, or equivalent solutions.
Experience working in government, defense, or other highly regulated environments is highly preferred.
Experience implementing or supporting CI/CD pipelines and DevOps practices.
Proficiency in infrastructure automation using Ansible, Shell scripting, Python, or similar automation tools.
Good understanding of system integration patterns, including REST APIs, message queues, and event-driven architectures.
Familiarity with cybersecurity standards, vulnerability management, and regulatory compliance requirements.
Strong knowledge of Platform-as-a-Service (PaaS), Container-as-a-Service (CaaS), container orchestration, COTS platforms, and operating system hardening.
Excellent analytical, troubleshooting, and problem-solving skills.
Strong communication and stakeholder management abilities, with experience working across cross‑functional technical teams.
Lead and mentor a team of 4–6 Infrastructure Engineers.
Foster a culture of operational excellence, continuous improvement, knowledge sharing, and technical development.
Coordinate team workload, provide technical guidance, and ensure adherence to operational processes and service level agreements (SLAs).