Principal Cloud Data Reliability Engineer

Oracle Corporation

Seattle (WA)

On-site

USD 140,000 - 190,000

Full time

7 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Global scale projects
Distributed systems exposure
Reliability-focused culture
Career growth opportunities
Flexible work arrangements
Competitive compensation and benefits

Job summary

Oracle Corporation is seeking a Principal Site Reliability Engineer to own operational health, reliability, and continuous improvement of OCI GoldenGate and related services. You will ensure production cloud services run securely, with automation driving efficiency and consistent deployments.

The role requires leading incident response, building observability, and partnering with software teams to scale globally. A TS clearance eligibility and 7+ years of experience are expected.

Qualifications

  • 7+ years production engineering experience (distributed services or cloud platform experience preferred).
  • Automation experience and scripting (bash, Python, etc.).
  • Solid understanding of Software Engineering and Computer Science principles.
  • Solid foundation in Linux Systems Engineering.
  • Solid understanding of Networking, Security and Storage.
  • Experience with Docker.
  • Understanding Infrastructure as Code concept.
  • Experience with Terraform and/or Ansible.
  • Experience building fully automated CI/CD pipelines (Jenkins, Hudson, TeamCity).
  • Ability to formulate meaningful KPIs and business metrics.
  • Experience with cloud services at major providers (AWS, Azure, GCP or OCI) is a big plus.

Responsibilities

  • Operate and support production cloud services that power critical customer migration and data replication workloads.
  • Monitor service health, investigate incidents, and lead troubleshooting efforts for complex production issues.
  • Participate in incident response, root cause analysis, and post-incident reviews with corrective actions.
  • Act as escalation point for complex issues not yet documented as SOPs.
  • Partner with software development teams to improve service architecture and operational readiness.
  • Design and implement automation to reduce operational overhead and improve efficiency.
  • Build and maintain observability solutions: monitoring, alerting, logging, dashboards, metrics.
  • Contribute to capacity planning, performance analysis, disaster recovery readiness, and resilience initiatives.
  • Support services across Oracle Cloud environments; leverage IaC and CI/CD practices.

Skills

Automation
SRE principles
Distributed systems
Networking
Linux systems engineering
Security and storage

Education

Bachelor's degree or equivalent

Tools

Docker
Terraform
Ansible
Jenkins
Hudson
TeamCity

Job description

Oracle Corporation is seeking a Principal Site Reliability Engineer to own operational health, reliability, and continuous improvement of OCI GoldenGate and related services. You will ensure production cloud services run securely, with automation driving efficiency and consistent deployments.

The role requires leading incident response, building observability, and partnering with software teams to scale globally. A TS clearance eligibility and 7+ years of experience are expected.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Principal Cloud Data Reliability Engineer
Principal Cloud Data Reliability Engineer

Oracle • Seattle (WA)

On-site
USD 115,000 - 235,000
Competitive compensation
Comprehensive benefits
Flexible work arrangements
Senior Cloud Reliability Engineer – Data Systems
Senior Cloud Reliability Engineer – Data Systems

Oracle • Herndon (VA)

On-site
USD 115,000 - 235,000
Medical, dental, and vision insurance
Paid time off
401(k) with company match
+3
Principal Data Systems Software Engineer - SRE
Principal Data Systems Software Engineer - SRE

Oracle Corporation • Seattle (WA)

On-site
USD 140,000 - 190,000
Global scale projects
Distributed systems exposure
Reliability-focused culture
+3
Senior Cloud Data Systems Engineer - Automation & Reliability
Senior Cloud Data Systems Engineer - Automation & Reliability

Ll Oefentherapie • Seattle (WA)

On-site
USD 150,000 - 210,000
Senior Principal Cloud Reliability Engineer - Databases
Senior Principal Cloud Reliability Engineer - Databases

Oracle • Santa Clara (CA)

On-site
USD 135,000 - 306,000
Medical, dental, and vision insurance
401(k) Retirement plan with company m|
Paid time off
+3
Principal Data Systems Software Engineer - SRE
Principal Data Systems Software Engineer - SRE

Ll Oefentherapie • Seattle (WA), Herndon (VA)

On-site
USD 120,000 - 180,000
Top Secret clearance required
Senior Site Reliability Engineer - Scalable Cloud & Automation
Senior Site Reliability Engineer - Scalable Cloud & Automation

Oracle • North Carolina

On-site
USD 84,000 - 210,000
Principal Data Systems Software Engineer
Principal Data Systems Software Engineer

Ll Oefentherapie • Seattle (WA)

On-site
USD 150,000 - 210,000
Senior Data Center Reliability & Operations Engineer
Senior Data Center Reliability & Operations Engineer

Oracle • Ashburn (VA)

On-site
USD 102,000 - 210,000
Medical, dental, and vision insurance
401(k) with company match
Paid time off
Senior Principal Reliability & Design Quality Engineer
Senior Principal Reliability & Design Quality Engineer

Oracle • United States

On-site
USD 146,000 - 306,000
Medical, dental, and vision insurance
401(k) Savings plan with company match
Paid time off