Senior SRE Lead: Cloud Reliability & Automation

Oracle

Vienna (VA)

On-site

USD 96,300 - 264,100

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical, dental, vision insurance
401(k) with company match
Paid time off and holidays
Parental leave

Job summary

Oracle is seeking a seasoned Site Reliability Engineer to help manage a cloud platform and mission‑critical apps. You will design resilient systems, implement IaC with Terraform, and lead automation efforts across observability, incident response, and capacity planning.

You will collaborate with cloud, security, and development teams to ensure high availability and secure, scalable deployments, while owning proactive health reporting and on‑call leadership.

Qualifications

  • Bachelor's degree in CS/IT/Engineering or equivalent experience.
  • 8+ years of experience in Site Reliability Engineering, DevOps, Cloud or Systems Engineering.
  • Hands-on experience with Oracle Cloud Infrastructure (OCI) or another major cloud provider.
  • Experience with Kubernetes, Docker, and container orchestration.

Responsibilities

  • Maintain high availability, reliability, and performance of enterprise applications and cloud infrastructure.
  • Design and implement automation to reduce manual operational effort and improve deployment consistency.
  • Develop Infrastructure as Code (IaC) using Terraform and automation scripts using Python, Bash, or similar languages.
  • Build and maintain CI/CD pipelines to support automated deployments and release management.
  • Monitor applications and infrastructure using observability platforms, including metrics, logs, traces, and alerting.
  • Define and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets.
  • Participate in on-call rotations and respond to production incidents with urgency and professionalism.
  • Perform capacity planning, performance tuning, and scalability assessments.
  • Collaborate with development teams to improve application resiliency and fault tolerance.
  • Provide guidance on monitoring, and document system health and performance.

Skills

SRE
OCI
Kubernetes
Docker
Terraform
Python
Bash
Linux
CI/CD
Observability
Incident Mgmt
RCA
Automation
Capacity Planning
Performance Tuning
Networking
Cloud Security
High Availability
Disaster Recovery
DevOps
Agile

Education

Bachelor's degree in Computer Science, Information Technology, Engineering, or equivalent experience

Tools

Jenkins
GitHub Actions
GitLab CI
Azure DevOps

Job description

Oracle is seeking a seasoned Site Reliability Engineer to help manage a cloud platform and mission‑critical apps. You will design resilient systems, implement IaC with Terraform, and lead automation efforts across observability, incident response, and capacity planning.

You will collaborate with cloud, security, and development teams to ensure high availability and secure, scalable deployments, while owning proactive health reporting and on‑call leadership.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer – Cloud Reliability & Automation
Senior Site Reliability Engineer – Cloud Reliability & Automation

Ll Oefentherapie • Reston (VA)

On-site
USD 85,000 - 210,000
Medical, dental, and vision insurance
401(k) with company match
Paid time off
+2
Senior SRE — Cloud Reliability, Automation & Equity
Senior SRE — Cloud Reliability, Automation & Equity

Oracle • United States

On-site
USD 81,000 - 187,000
Medical, dental and vision insurance
401(k) Savings and Investment Plan
Paid time off and holidays
+3
Senior Principal SRE Lead — Cloud Reliability & Automation
Senior Principal SRE Lead — Cloud Reliability & Automation

Oracle • United States

On-site
USD 105,000 - 264,000
Senior SRE Manager: Reliability, Automation & Scale
Senior SRE Manager: Reliability, Automation & Scale

Oracle • San Juan (PR)

On-site
USD 150,000 - 190,000
Senior SRE - Cloud Infra, Automation & 24/7 Ops
Senior SRE - Cloud Infra, Automation & 24/7 Ops

Oracle • Reston (VA)

On-site
USD 85,000 - 210,000
Medical, dental, and vision insurance
Paid time off
401(k) Savings Plan
Senior Cloud Site Reliability Engineer — Automation
Senior Cloud Site Reliability Engineer — Automation

Ll Oefentherapie • Richmond (VA)

On-site
USD 120,000 - 180,000
Senior Site Reliability Engineering Lead
Senior Site Reliability Engineering Lead

Oracle • Frankfort (KY)

On-site
USD 122,000 - 264,000
Medical, dental, and vision insurance
401(k) Savings and Investment Plan
Paid time off
Senior Cloud SRE — 24/7 Automation & AI Infra
Senior Cloud SRE — 24/7 Automation & AI Infra

Oracle • Austin (TX)

On-site
USD 84,000 - 210,000
Cloud SRE Engineer: Reliability, Scale & Automation
Cloud SRE Engineer: Reliability, Scale & Automation

Oracle • San Juan (PR)

On-site
USD 74,000 - 148,000
Health insurance
Employee stock purchase plan
Paid time off
Senior Site Reliability & Infrastructure Leader
Senior Site Reliability & Infrastructure Leader

Oracle • United States

On-site
USD 121,500 - 264,100
Medical insurance
Dental insurance
Vision insurance
+7