Get more replies from employers
Send a job-specific resume in minutes.
Unosquare seeks a Senior Site Reliability Engineer to apply Google-inspired SRE practices to hybrid systems running OCI and IBM Mainframe with z/OS 2.5. You will automate operations, define SLOs, reduce toil, and support CA-IDEAL online and COBOL batch apps.
The role blends legacy mainframe with cloud-native modernization and data replication or API orchestration on OCI. Ideal candidates will have 6+ years in SRE or mainframe operations, strong Python/Go skills, and expertise in z/OS 2.5,
We are seeking a Senior Site Reliability Engineer who applies Google-inspired SRE principles to ensure the reliability of hybrid systems spanning Oracle Cloud Infrastructure (OCI) and IBM Mainframe environments running z/OS 2.5.
You'll focus on production reliability as a software engineering discipline: automating operations, defining SLOs, reducing toil, and handling incidents, while providing specialized support for CA-IDEAL online applications and COBOL batch applications. This role is perfect for a mainframe expert eager to blend legacy systems with cloud-native practices, enabling seamless integration, modernization, or migration of mainframe workloads to OCI or other SaaS cloud (e.g., via data replication, API orchestration, or hybrid architectures leveraging OCI's compute, storage, and networking).
Embrace blameless postmortems, error budgets, and automation to keep critical business processes running at scale.
There's an opportunity to grow experience into SRE practices running Oracle EBS on OCI.
Own the end-to-end reliability, performance, availability, and scalability of integrated OCI and IBM Mainframe systems
Establish and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets for mainframe-dependent services and OCI workloads
Develop automation tools and scripts to manage mainframe operations on z/OS 2.5, using languages like Python, Go, REXX, or JCL-integrated automation
Participate in on-call rotation for incident response, troubleshooting, root-cause analysis, and blameless postmortems across hybrid environments
Eliminate toil by automating repetitive mainframe tasks: job scheduling, batch monitoring, error handling, and system maintenance
Provide expert support for IBM Mainframe components:Administer z/OS 2.5 environments, including JES2/JES3, SMF, RACF security, WLM, and system performance tuning
Manage CA-IDEAL online applications: development, deployment, runtime support, integration with CA Datacom/DB, screen mapping, and transaction processing (e.g., under CICS or IMS)
Handle COBOL batch applications: JCL creation/submission, debugging, VSAM/ISAM file management, DB2 or IDMS database interactions, and batch optimization
Ensure high availability and disaster recovery for mainframe apps (e.g., using GDPS, Parallel Sysplex, or integration with OCI DR solutions)
Troubleshoot issues in hybrid setups: OCI services (Compute, Networking, Storage, Databases) interfacing with mainframe via connectors like OCI GoldenGate, API Gateway, or custom middleware
Conduct capacity planning, cost optimization, and performance tuning for mainframe-OCI integrations, including data synchronization and workload offloading
Promote SRE culture across teams, mentoring on mainframe reliability best practices