AVP, SRE Engineer (Control-M), Technology Group

GIC Private Limited

Singapore

On-site

SGD 180,000 - 240,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

GIC Private Limited is seeking an AVP, SRE Engineer (Control-M) to own the design, deployment, and reliability of the enterprise batch scheduling estate. You will partner with platform, cloud, and operations teams to embed reliability, automation, and observability across workloads.

The role requires deep Control-M expertise, strong automation skills, and experience with on‑prem and cloud environments (AWS/Azure).

Qualifications

  • Bachelor’s or Master’s degree required in CS/Engineering or related field.
  • 7+ years in IT operations, Infrastructure, or SRE with emphasis on Control-M/batch scheduling.
  • Hands-on with Control-M enterprise manager/server/agents and HA/DR configurations.

Responsibilities

  • Own end-to-end Design, implementation, and reliability of Control-M batch estate.
  • Automate recovery, observability, and self-healing workflows using Python, Shell, Ansible, Terraform.
  • Integrate Control-M with CI/CD and Jobs-as-Code via Automation API/ctm CLI.
  • Lead upgrades, capacity planning, and governance for job definitions.
  • Partner with infra, cloud, and app teams to onboard workloads and ensure resilience.

Skills

Control-M
SRE
Python
Shell scripting
Ansible
Terraform
Linux/Unix
AWS
Azure
Oracle/SQL

Education

Bachelor’s or Master’s degree in Computer Science or Engineering

Tools

Control-M Automation API
MFT
Datadog
Splunk
ServiceNow

Job description

Select how often (in days) to receive an alert: Create Alert

AVP, SRE Engineer (Control-M), Technology Group

Location: Singapore, SG

Job Function: Technology Group

Job Type: Permanent

GIC is one of the world’s largest sovereign wealth funds. With over 2,000 employees across 11 locations around the world, we invest in more than 40 countries globally across asset classes and businesses. Working at GIC gives you exposure to an extraordinary network of the world’s industry leaders. As a leading global long-term investor, we Work at the Point of Impact for Singapore’s financial future, and the communities we invest in worldwide.

Technology Group We experiment, design, and lead a 24×7 global business where we support core capabilities in asset management, trading, investment operations, and risk management. We deliver secure, reliable, and integrated solutions, and provide insights on new, and emerging technologies.

Infrastructure & Cybersecurity Resilience (ICR) We design, build, and secure the technology foundations that power GIC’s global investment operations. We aim to deliver resilient, scalable, and secure infrastructure that empowers our people and businesses to perform securely, efficiently, and effectively.

What impact will you make in this role? As an SRE Engineer, you will be the hands‑on subject matter expert (SME) responsible for the design, implementation, optimization, and reliability of the enterprise workload automation estate. You will own the Control‑M platform end to end – ensuring that mission‑critical batch schedules across infrastructure and application domains run reliably, recover gracefully, and meet stringent operational resilience and regulatory standards. This is an individual contributor and deep hands‑on role.

You will bring authoritative expertise in Control‑M and batch scheduling products, partnering closely with application, infrastructure, cloud, and operations teams to embed reliability, automation, and observability into every workflow. You will act as the highest escalation point for complex scheduling incidents and drive the modernization of the batch estate toward self‑healing, cloud‑native, and event‑driven automation.

What will you do as a SRE Engineer (Control-M)?

Platform Engineering & Architecture

  • Architect, deploy, and maintain enterprise Control‑M / BMC Helix Control‑M environments (Enterprise Manager, Server, Agents, and the high‑availability/failover topology) aligned with operational resilience mandates such as MAS TRM, DORA, and APRA CPS 230.
  • Design scalable, resilient batch scheduling patterns and reusable job templates, calendars, and folder structures that standardize how workloads are onboarded across the enterprise.
  • Lead platform upgrades, patching, version migrations, and high‑availability/disaster‑recovery configurations with minimal disruption to production schedules.
  • Act as the Control‑M SME for architecture reviews, capacity planning, and performance tuning of the scheduling estate.
  • Define and enforce job‑definition, naming, and governance standards—including retention, audit logging, and change control for the batch environment.

Reliability Engineering & Automation

  • Champion SRE frameworks and reliability practices for mission‑critical batch workloads, defining SLAs, error budgets, and service‑health indicators for scheduled jobs.
  • Design and automate self‑healing workflows, intelligent job recovery, and auto‑remediation using Control‑M capabilities together with Python, Shell, Ansible, and Terraform.
  • Integrate Control‑M with CI/CD pipelines using “Jobs‑as‑Code” (Control‑M Automation API / ctm CLI) to enable version‑controlled, auditable, and automated deployment of job definitions.
  • Reduce operational toil by replacing manual interventions with automated, observable, and repeatable scheduling workflows.
  • Conduct resilience, failover, and capacity testing aligned with business continuity and disaster‑recovery standards.

Advanced Troubleshooting & Incident Management

  • Serve as the highest escalation point for complex batch and scheduling incidents, performing deep‑dive root cause analysis across job dependencies, infrastructure, and integrated applications.
  • Drive blameless post‑incident reviews and codify lessons learned into automation, monitoring, and runbook improvements.
  • Proactively identify and resolve SLA breaches, long‑running jobs, and dependency bottlenecks before they impact downstream business processes.

Integration, Observability & Operational Excellence

  • Integrate Control‑M with ITSM/incident management (e.g., ServiceNow), monitoring, and AIOps platforms to enable predictive alerting and anomaly detection on batch workloads.
  • Engineer batch scheduling for hybrid and cloud‑native workloads across AWS and Azure, orchestrating Managed File Transfer (MFT), data pipelines, ETL/ELT, and application workflows.
  • Develop and maintain executive dashboards and reports showcasing batch availability, SLA adherence, job‑failure trends, and operational‑risk indicators.
  • Drive measurable reduction in incident recurrence, MTTR, and manual intervention through automation‑led observability.
  • Partner with application, infrastructure, cloud, DevOps, and operations teams to onboard new workloads and embed reliability principles into the delivery lifecycle.
  • Mentor junior engineers and operations staff on Control‑M best practices, troubleshooting, and automation; create documentation, runbooks, and training materials.
  • Evaluate emerging workload‑automation and orchestration technologies (e.g., cloud‑native schedulers, Apache Airflow, event‑driven orchestration) to modernize the batch estate.

What makes you a successful candidate?

  • Bachelor’s or Master’s degree in Computer Science, Engineering, or a related discipline.
  • 6‑9+ years of experience in IT operations, infrastructure, or SRE roles, with at least 7+ years specializing hands‑on in Control‑M / batch scheduling engineering, ideally in financial or other regulated environments.
  • Deep, hands‑on expertise as a Control‑M SME—administration, architecture, upgrades, and performance tuning of Enterprise Manager, Server, and Agents.
    Proven hands‑on expertise in:
    – Scheduling & Workload Automation: Control‑M / BMC Helix Control‑M (must), Control‑M Automation API (Jobs‑as‑Code), MFT, Managed File Transfer, and exposure to alternatives such as Autosys, TWS/IWS, or Apache Airflow.
    – Automation / IaC: Python, Shell scripting, Ansible, Terraform, and CI/CD tooling (Git, Jenkins, Azure DevOps).
    – Operating Systems & Databases: Strong Linux/Unix and Windows administration; working knowledge of Oracle, SQL Server, or PostgreSQL underpinning the scheduling estate.
    – Cloud & Integration: AWS and Azure batch/orchestration services, ServiceNow, and monitoring/observability tooling (e.g., Datadog, Splunk).
  • Deep understanding of SRE principles, service‑health modelling, error budgets, SLA management, and auto‑remediation design for batch workloads.
  • Strong analytical and troubleshooting skills, with the ability to perform deep‑dive investigations across complex job dependencies and develop long‑term preventive solutions.
  • Solid understanding of job‑dependency design, calendars, cyclic and event‑driven scheduling, SLA/critical‑path management, and high‑availability/disaster‑recovery configurations.
  • Familiarity with financial‑sector operational resilience frameworks, regulatory compliance, change governance, and audit requirements.
  • Working understanding of security foundations—IAM, role‑based access control, credential management, and data protection—as applied to the scheduling platform.
  • Soft skills & mindset: Systems thinking with the ability to connect scheduling, infrastructure, and application workflows; ownership and accountability for production stability; and strong collaboration across operations, application, and engineering teams.

Work at the Point of Impact We need to be forward‑looking to attract the right people to help us become the Leading Global Long‑term Investor. Join our ambitious, agile, and diverse teams—be empowered to push boundaries and pursue innovative ideas, share your views, and be heard. Be anchored on our PRIME Values: Prudence, Respect, Integrity, Merit and Excellence, which guides us in how we make our day‑to‑day decisions. We strive to inspire. To make an impact.

GIC is a Great Place to Work

At GIC, our offices are vibrant hubs for ideation, professional growth, and interpersonal connection. At the same time, we believe that flexibility allows us to do our best work and be our best selves. Thus, our teams come into the office four days per week to harness the benefits of in‑person collaboration, but have the flexibility to choose which days they work from home and adjust this arrangement as situational needs arise.

GIC is an equal opportunity employer As an employer, we passionately believe every individual brings with them unique diversity of thought and perspectives to meaningfully enrich perspectives of GIC teams to drive competitive performance. An inclusive environment yields exceptional contribution.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AVP, SRE Engineer (Control-M), Technology Group
AVP, SRE Engineer (Control-M), Technology Group

GIC • Singapore

On-site
SGD 120,000 - 160,000
AVP/VP, Network Automation and Reliability Engineer, Technology Group
AVP/VP, Network Automation and Reliability Engineer, Technology Group

GIC Private Limited • Singapore

Hybrid
SGD 180,000 - 280,000
AVP, Ansible Automation Platform & AIOps Developer, Technology Group
AVP, Ansible Automation Platform & AIOps Developer, Technology Group

GIC Private Limited • Singapore

Hybrid
SGD 120,000 - 180,000
Flexible work arrangement
Office four days per week
VP/SVP, Production Support Lead (SAT), Technology Group
VP/SVP, Production Support Lead (SAT), Technology Group

GIC Private Limited • Singapore

On-site
SGD 180,000 - 240,000
Associate/AVP, Observability & SRE Engineering, Technology Group
Associate/AVP, Observability & SRE Engineering, Technology Group

GIC • Singapore

Hybrid
SGD 150,000 - 195,000
AVP/VP, Network Engineer, Technology Group
AVP/VP, Network Engineer, Technology Group

GIC Private Limited • Singapore

On-site
SGD 80,000 - 120,000
Flexible work arrangements
Professional growth opportunities
AVP/VP, Product Engineer (DevSecOps), Technology Group
AVP/VP, Product Engineer (DevSecOps), Technology Group

GIC Private Limited • Singapore

On-site
SGD 120,000 - 180,000
AVP/VP, SIEM & SRE Engineering, Technology Group
AVP/VP, SIEM & SRE Engineering, Technology Group

GIC • Singapore

Hybrid
SGD 120,000 - 160,000
SVP, Production Support Lead, Technology Group
SVP, Production Support Lead, Technology Group

GIC Private Limited • Singapore

On-site
SGD 120,000 - 160,000
AVP/VP, Incident and Problem Management, Technology Group
AVP/VP, Incident and Problem Management, Technology Group

GIC Private Limited • Singapore

Hybrid
SGD 120,000 - 180,000