Senior Mainframe Systems Programmer - Site Reliability Engineering

Ensono

Pune District

On-site

INR 1,800,000 - 2,400,000

Full time

3 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Ensono in Pune, India, seeks a visionary Mainframe Site Reliability Engineer (SRE) to redefine reliability and automation for z/OS systems. You will drive innovations in observability, AI‑driven operations, and DevOps, transforming legacy workflows into self‑healing, scalable platforms.

You will lead automation across CICS/Db2/IMS, implement AI‑driven monitoring, and integrate CI/CD pipelines for mainframe apps, with a focus on SLOs and resilience.

Qualifications

  • Bachelor’s degree in Computer Science or a related field.
  • xx+ years in z/OS system programming, performance tuning, or infrastructure support.
  • Proficiency in JCL, REXX, Python, and mainframe automation tools.
  • Hands-on experience with Zowe, Ansible, Git, and CI/CD pipelines.

Responsibilities

  • Design IaC solutions using Ansible and z/OSMF to automate provisioning and recovery.
  • Develop self-healing workflows for CICS/Db2/IMS to auto-resolve incidents.
  • Implement AI-driven observability with tools like IBM Watson AIOps and Splunk ITSI.
  • Build dashboards with Grafana/Prometheus for mainframe metrics.
  • Streamline COBOL/PL/I delivery with DBB and UCD in CI/CD pipelines.
  • Lead blameless postmortems and reduce MTTR.
  • Mentor teams on SRE principles and DevOps practices.
  • Optimize batch workloads with dynamic resource management.
  • Ensure security hardening (RACF, TLS) and runbooks.

Skills

z/OS expert
JCL & REXX
Python
Mainframe automation
Ansible
Git CI/CD
SRE basics
Zowe automation

Education

Bachelor's in CS/Engineering

Tools

Zowe
Git
Jenkins
UrbanCode Deploy
IBM Z Automation

Job description

About Us

Ensono is a hybrid IT services provider that sees a world in transformation. We understand the world of business has become more complex, layered, and interdependent because we’re enabling the IT infrastructure for that transformation. Our clients are some of the most innovative and forward-thinking companies in the world and we help keep their businesses thriving. With nearly 50 years of experience, we optimize and modernize IT infrastructure by amplifying the power of mainframes and mid-range servers with the agility of cloud. Check us out at www.ensono.com

Job Role

We are seeking a visionary Mainframe Site Reliability Engineer (SRE) to redefine the reliability, automation, and efficiency of our mission‑critical z/OS systems. This role combines deep mainframe expertise with cutting‑edge SRE practices, focusing on innovations in observability, AI‑driven operations, and DevOps integration to transform legacy workflows into modern, self‑healing systems. You will drive initiatives to eliminate manual toil, optimize performance, and ensure the platform’s resilience aligns with business‑critical service level objectives (SLOs).

Job Responsibilities
1.SRE-Centric Innovation & Automation
Automation Engineering
  • Design and deploy Infrastructure-as-Code (IaC) solutions using Ansible, Zowe CLI, and z/OSMF workflows to automate system provisioning, configuration management, and recovery processes.
  • Develop self‑healing workflows for critical subsystems (CICS, Db2, IMS) to auto‑resolve incidents like JVM failures or transaction bottlenecks.
  • Convert legacy operational scripts (REXX, NCL) into modern, version‑controlled pipelines integrated with Git and CI/CD tools like Jenkins.
AI-Driven Observability
  • Implement predictive analytics tools (e.g., IBM Watson AIOps, Splunk ITSI) to detect anomalies in system metrics, logs, and message queues.
  • Build dashboards using Grafana or Prometheus to visualize the Four Golden Signals (latency, traffic, errors, saturation) across mainframe workloads.
  • Centralize alert management to reduce noise and prioritize actionable alerts using AI‑driven correlation.
DevOps Integration & Modernization
CI/CD for Mainframe
  • Streamline software delivery pipelines for COBOL/PL/I applications using IBM Dependency-Based Build (DBB) and UrbanCode Deploy (UCD).
  • Integrate mainframe SDLC processes with enterprise Git repositories (GitHub, GitLab) to enable collaborative development and audit trails.
  • Enable automated testing and phased rollouts for z/OS middleware updates.
Performance & Capacity Engineering
  • Optimize CPU/MIPS utilization through runtime tuning (e.g., CICS Threadsafe, AT‑TLS offloading) to reduce software licensing costs.
  • Forecast capacity demands using historical SMF/RMF data and propose dynamic hardware scaling strategies.
  • Conduct load testing for batch and OLTP workloads to validate system limits and error budgets.
Incident Management & Reliability
  • Lead blameless postmortems for critical incidents, focusing on root cause analysis (RCA) and preventive actions (e.g., monitoring gaps, automation fixes).
  • Reduce MTTR by implementing automated incident response playbooks (e.g., auto‑restart failed subsystems, reroute traffic).
  • Maintain 24/7 operational readiness through on‑call rotations and cross‑training in z/OS, CICS, Db2, and storage management.
Platform Hardening & Knowledge Sharing
  • Enforce security best practices (RACF, TLS) and vulnerability remediation for z/OS and middleware.
  • Develop reusable workbooks and runbooks to document system configurations, troubleshooting steps, and automation workflows.
  • Mentor teams on SRE principles, fostering a T‑shaped skill model (deep mainframe + DevOps/Agile practices).
Batch Optimization & Resource Management
  • Design dynamic resource allocation strategies (e.g., WLM policies, enclaves) to prioritize critical batch jobs and minimize contention for CPU, memory, and I/O resources.
  • Implement parallel processing (e.g., multi‑task JCL, SYSAFF routing) to reduce runtime and avoid bottlenecks in long‑running batch cycles.
  • Streamline job dependencies using graph‑based scheduling tools (e.g., IWS, CA7, Control‑M) to eliminate idle wait times between interdependent jobs.
Proactive Batch Health Monitoring
  • Develop automated checks for batch job SLAs, including real‑time alerts for delays, resource starvation, or dataset contention.
  • Integrate predictive analytics (e.g., historical SMF data analysis) to forecast and mitigate delays caused by seasonal peaks or data volume spikes.

---

Required Skills
Technical Expertise
  • xx+ years in z/OS system programming, performance tuning, or infrastructure support.
  • Proficiency in JCL, REXX, Python, and mainframe automation tools (IBM Z System Automation, Broadcom OPS/MVS).
  • Hands‑on experience with Zowe, Ansible, Git, and CI/CD pipelines.
  • Mastery of SRE tenets: SLOs/SLIs, error budgets, and Infrastructure-as-Code (IaC).
Innovation Focus
  • Proven track record in implementing AI/ML‑driven monitoring or auto‑remediation for mainframe environments.
  • Experience modernizing legacy workflows (e.g., replacing CA Endevor with Git‑based SDLC).
Soft Skills
  • Ability to lead cross‑functional teams during high‑severity incidents.
  • Strong communication to align technical execution with business objectives.
Education
  • Bachelor’s degree in Computer Science, Engineering, or related field.

---

Preferred Qualifications
  • Experience with AI‑Driven Automation platforms (e.g. AMELIA AIOps) to standardize and migrate legacy workflows, integrate with event management systems (e.g., BigPanda), and orchestrate ITIL processes (Incident, changes) via ServiceNow
  • Certifications: IBM z/OS System Programming, Broadcom Mainframe SRE, or Hashicorp Terraform.
  • Familiarity with Zowe Desktop for modern IDE‑driven development or Dynatrace APM for CICS/Db2 monitoring.
  • Knowledge of mainframe open‑source ecosystems (Zowe, Feilong) or hybrid‑cloud integrations.

Shift Timing- 1:30 PM to 10:30 PM

We are an equal opportunity employer. All qualified applicants will be considered for employment without regard to caste, colour, creed, religion, gender, gender identity, sexual orientation, age, disability, HIV status, or any other status protected by law. Candidates with disabilities who require accommodations during the recruitment process are encouraged to contact our Talent Acquisition team to place a request.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior z/OS Systems Programmer
Senior z/OS Systems Programmer

Ensono • Hyderabad

On-site
INR 2,500,000 - 4,500,000
Expert Software Engineer Mainframe AMS
Expert Software Engineer Mainframe AMS

Ensono • Chennai District

On-site
INR 1,200,000 - 1,800,000
Senior z/OS Unix & zLinux Integration Engineer
Senior z/OS Unix & zLinux Integration Engineer

Ensono • Hyderabad

On-site
INR 11,483,000 - 17,225,000
Principal System Administrator (Z/VM , Z/OS, ZLinux)
Principal System Administrator (Z/VM , Z/OS, ZLinux)

Oracle • Bengaluru

On-site
INR 1,800,000 - 2,500,000
Infrastructure Senior Technology Analyst - Assistant Vice President
Infrastructure Senior Technology Analyst - Assistant Vice President

Citi • Chennai District

On-site
INR 2,500,000 - 4,500,000
Mainframe Systems Programmer - Network
Mainframe Systems Programmer - Network

Ensono • Pune District

On-site
INR 2,500,000 - 4,000,000
Z/OS System Administrator
Z/OS System Administrator

AagatiServe Pvt Ltd • India

On-site
INR 1,500,000 - 2,100,000
Mainframe z/OS Systems Programmer
Mainframe z/OS Systems Programmer

OP • India

On-site
INR 3,500,000 - 7,500,000
Health Insurance
Accident Insurance
Senior Mainframe Systems Programmer - CICS/MQ
Senior Mainframe Systems Programmer - CICS/MQ

Ensono • Hyderabad

On-site
INR 1,800,000 - 2,800,000
Zos System Programmer
Zos System Programmer

AagatiServe Pvt Ltd • Gurgaon

On-site
INR 2,500,000 - 4,000,000