Senior Software Engineer (Cloud Infrastructure / Site Reliability Engineering)

Oscar Health

Los Angeles (CA)

On-site

USD 180,000 - 240,000

Full time

6 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Oscar Health is seeking a Staff/Senior Staff Engineer to lead critical infrastructure initiatives across Core Technology. You will guide cross-team deliveries, mentor engineers, and shape the technical roadmap while ensuring system resilience and security at scale.

You will own medium to large features, collaborate with product and design, and advocate best practices in cloud-native environments using AWS/GCP, Terraform, and Kubernetes.

Qualifications

  • 6+ years of professional software engineering experience.
  • Experience leading cross-pod or cross-company deliverables.
  • Experience mentoring junior engineers and setting coding standards.
  • Strong CS fundamentals and scalable systems design.
  • Education: BS in CS or related field.

Responsibilities

  • Lead planning, execution and release of complex technical projects across multiple teams outside Core Technology.
  • Collaborate with partners, product managers and designers to solve problems.
  • Lead and mentor engineers to improve technology and practices.
  • Own large features or infrastructure capabilities within or across domains.
  • Facilitate cross-team collaboration and mitigate risks to deliver on time.
  • Drive technical roadmap and influence product/process priorities.
  • Design resilient systems and reduce outage impact.
  • Develop software to minimize maintenance during failures.
  • Define and guide Service-Level Objectives (SLOs).
  • Ensure compliance with applicable laws and regulations.

Skills

System design
Distributed systems
Cloud-native security
SRE practices
Mentoring
Leadership
Automation
Incident management
CI/CD

Education

B.S. in Computer Science
Equivalent professional experience

Tools

Kubernetes
ArgoCD
Terraform
Istio
GitHub Actions
Prometheus
Grafana
AWS
GCP

Job description

  • Our Core Technology teams build and maintain the foundational platform upon which all Oscar engineering is built
  • We are responsible for architecting a world-class, resilient ecosystem using a modern stack centered on AWS/GCP, Terraform, and Kubernetes with developer focused tooling and CI/CD
  • Our mission is to provide an automated, self-service infrastructure that empowers our engineering organization to move fast without sacrificing security or stability
  • You will report into a Staff/Senior Staff Engineer
  • Become the expert on your team’s business and technical domains such as DevOps, site reliability, and cloud best practices
  • Lead the planning, execution and release of complex technical projects across multiple teams outside of Core Technology
  • Work with partners, product managers, and designers to solve challenging problems
  • Lead and mentor engineers on the team to improve technology and apply best practices
  • Independently responsible for large or complex technology capabilities (set of components or services) within their team’s domain or spanning multiple domains
  • Facilitates, encourages, and enhances cross-team execution and collaboration; knows when cross-team projects are at risk and actively mitigates risk to deliver on time
  • Prolific contributor to the objectives of their functional group, as well as organization-wide projects
  • Drives prioritization of technical roadmap and influences prioritization of product roadmap and process enhancements within their team
  • Actively identifies and reduces failure domains, designs and builds resilient systems, and strives to reduce adverse effects of an outage
  • Builds software to minimize effort and business impact during maintenance and failures
  • Guides the development of Service-Level Objectives (SLOs) for systems they are responsible for
  • Own medium to large features or infrastructure projects from technical design through completion
  • Compliance with all applicable laws and regulations
  • Other duties as assigned

6+ years of professional software engineering experience, working with a variety of technologies, and have increasingly impactful accomplishmentsExperience as a major contributor cross-pod or cross-company deliverablesExperience leading technical contributions, improving the quality of what your teams create, and are excited to build fault-tolerant, and scalable software systemsExperience mentoring and training more junior engineersSets and enforces the standard for writing stable, correct, and maintainable codeDemonstrates expertise of the practical application of CS concepts within their teamEducation: B.S. in Computer Science, a related technical field, or equivalent high-level industry experienceSecurity & Networking: Knowledge of cloud-native security (IAM, VPC peering) and service mesh technologies like IstioProgramming: Understanding of at least one coding language that you are able to use to develop scripts and softwareObservability: Proficiency with monitoring using tools like Prometheus, Grafana, or similarCI/CD & Automation: Experience building robust deployment pipelines via GitHub ActionsSRE Discipline: Strong background in Site Reliability Engineering, including Service Level Objectives (SLOs), error budgets, and incident managementOrchestration & Delivery: Proven track record with Kubernetes and workflows using ArgoCDInfrastructure as Code: Advanced experience with Terraform or similar IaC tools to manage complex, multi-account structuresCloud Proficiency: Deep expertise in managing production environments within AWS or GCP at scale

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Software Engineer (Cloud Infrastructure / Site Reliability Engineering (SRE)
Senior Software Engineer (Cloud Infrastructure / Site Reliability Engineering (SRE)

Oscar Health • San Francisco (CA)

On-site
USD 180,000 - 260,000
Staff Fullstack Software Engineer
Staff Fullstack Software Engineer

Oscar Health • Los Angeles (CA)

On-site
USD 150,000 - 210,000
Senior Cloud SRE & Platforms Engineer (AWS/GCP)
Senior Cloud SRE & Platforms Engineer (AWS/GCP)

Oscar Health • Los Angeles (CA)

On-site
USD 180,000 - 240,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

Staffing Science • Arizona

On-site
USD 180,000 - 240,000
Senior Cloud Infra & SRE Engineer
Senior Cloud Infra & SRE Engineer

Oscar Health • San Francisco (CA)

On-site
USD 180,000 - 260,000
Principal Site Reliability Engineer
Principal Site Reliability Engineer

Gen Digital Inc. • United States

Remote
USD 180,000 - 240,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

theaccessgroup • United States

Hybrid
USD 140,000 - 180,000
Health insurance
Dental insurance
Vision insurance
+3
Senior Lead Site Reliability Engineer
Senior Lead Site Reliability Engineer

JPMorgan Chase & Co. • Jersey City (NJ)

On-site
USD 150,000 - 210,000
Principal Dev-Ops Architect
Principal Dev-Ops Architect

Zohorecruit • United States

Remote
USD 180,000 - 240,000
Senior Software Engineer, Cloud Infrastructure / SRE
Senior Software Engineer, Cloud Infrastructure / SRE

Namely • San Francisco (CA)

On-site
USD 181,000 - 237,000