Principal Cloud Data Reliability Engineer

Oracle

Seattle (WA)

On-site

USD 115,000 - 235,000

Full time

4 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Competitive compensation
Comprehensive benefits
Flexible work arrangements

Job summary

Oracle is seeking a Principal Site Reliability Engineer to own the operational health and reliability of OCI GoldenGate and OCI Database Migration Service. You will partner with software teams to improve architecture, reliability, and scalability across Oracle Cloud environments.

In this role you will design automation, build observability solutions, and drive continuous service improvement through incident response, post-incident reviews, and SOP development, enabling cloud-native data movement

Qualifications

  • US citizen with Top Secret security clearance and SCI eligibility.
  • 7+ years production engineering experience (distributed services or cloud platform preferred).
  • Automation scripting experience in Bash, Python, etc.
  • Solid understanding of Software Engineering and CS principles.
  • Strong Linux systems engineering foundation.
  • Solid understanding of Networking, Security and Storage.
  • Experience with Docker.
  • Understanding Infrastructure as Code concepts.
  • Experience with Terraform and/or Ansible.
  • Experience building automated CI/CD pipelines (Jenkins, Hudson, TeamCity, etc).
  • Ability to define meaningful KPIs and business metrics.
  • Experience with cloud services at AWS, Azure, GCP or OCI is a big plus.

Responsibilities

  • Operate and support production cloud services that power critical customer migration and data replication workloads.
  • Monitor service health, investigate incidents, and lead troubleshooting efforts for complex production issues.
  • Participate in incident response, root cause analysis, and post-incident reviews with corrective actions.
  • Act as escalation point for complex issues not documented as SOPs.
  • Partner with software teams to improve service architecture, reliability, and operational readiness.
  • Design and implement automation to reduce operational overhead and improve efficiency.
  • Build and maintain observability solutions: monitoring, alerting, logging, dashboards and metrics.
  • Contribute to capacity planning, performance analysis, disaster recovery readiness, and resilience.
  • Support services across Oracle Cloud commercial and sovereign environments.
  • Leverage IaC and CI/CD practices to improve deployment consistency and scalability.
  • Drive continuous service improvement through reviews and reliability initiatives.

Skills

Distributed systems
Automation
Scripting (bash, python)
Linux systems
Networking
Security
CI/CD
Cloud platforms
Observability

Tools

Docker
Terraform
Ansible

Job description

Oracle is seeking a Principal Site Reliability Engineer to own the operational health and reliability of OCI GoldenGate and OCI Database Migration Service. You will partner with software teams to improve architecture, reliability, and scalability across Oracle Cloud environments.

In this role you will design automation, build observability solutions, and drive continuous service improvement through incident response, post-incident reviews, and SOP development, enabling cloud-native data movement

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Cloud Reliability Engineer – Data Systems
Senior Cloud Reliability Engineer – Data Systems

Oracle • Herndon (VA)

On-site
USD 115,000 - 235,000
Medical, dental, and vision insurance
Paid time off
401(k) with company match
+3
Principal Cloud Data Reliability Engineer
Principal Cloud Data Reliability Engineer

Oracle Corporation • Seattle (WA)

On-site
USD 140,000 - 190,000
Global scale projects
Distributed systems exposure
Reliability-focused culture
+3
Senior Principal Cloud Reliability Engineer - Databases
Senior Principal Cloud Reliability Engineer - Databases

Oracle • Santa Clara (CA)

On-site
USD 135,000 - 306,000
Medical, dental, and vision insurance
401(k) Retirement plan with company m|
Paid time off
+3
Principal Data Systems Software Engineer - SRE
Principal Data Systems Software Engineer - SRE

Oracle Corporation • Seattle (WA)

On-site
USD 140,000 - 190,000
Global scale projects
Distributed systems exposure
Reliability-focused culture
+3
Senior Principal Reliability & Design Quality Engineer
Senior Principal Reliability & Design Quality Engineer

Oracle • United States

On-site
USD 146,000 - 306,000
Medical, dental, and vision insurance
401(k) Savings plan with company match
Paid time off
Senior Site Reliability Engineer - Scalable Cloud & Automation
Senior Site Reliability Engineer - Scalable Cloud & Automation

Oracle • North Carolina

On-site
USD 84,000 - 210,000
Principal Cloud DB Platform Engineer
Principal Cloud DB Platform Engineer

Oracle • United States

Remote
USD 180,000 - 260,000
Senior Cloud Infra Engineer - Scalable, Reliable Systems
Senior Cloud Infra Engineer - Scalable, Reliable Systems

Oracle Corporation • Nashville (TN)

On-site
USD 79,000 - 210,000
Medical insurance
Dental insurance
Vision insurance
+11
Senior Data Center Reliability & Operations Engineer
Senior Data Center Reliability & Operations Engineer

Oracle • Ashburn (VA)

On-site
USD 102,000 - 210,000
Medical, dental, and vision insurance
401(k) with company match
Paid time off
Lead Principal Cloud Infrastructure Architect
Lead Principal Cloud Infrastructure Architect

Oracle • United States

On-site
USD 146,000 - 306,000
Medical insurance
Dental insurance
Vision insurance
+4