Senior Site Reliability Engineer

Red Hat, LLC

Raleigh (NC)

On-site

USD 118,600 - 195,680

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Comprehensive medical, dental, and vision coverage
Flexible Spending Account
401(k) with employer match
Paid parental leave

Job summary

Red Hat, LLC is looking for a Senior Software Engineer to join the HCP Platform Engineering team. In this role, you will build and operate Red Hat OpenShift Service on AWS, contributing to upstream open-source projects.

With a strong focus on production systems, you'll be involved in coding and deploying critical features, participating in on-call rotations, and improving observability. Ideal candidates will have 4+ years of software engineering experience especially in Go and Kubernetes.

Qualifications

  • 4+ years of software engineering experience in cloud environments.
  • Strong proficiency in Go with production-quality code.
  • Deep hands-on Kubernetes experience for building and operating clusters.

Responsibilities

  • Contribute production-grade software to upstream open source projects.
  • Operate as a full-stack systems owner and participate in on-call rotations.
  • Design and evolve observability metrics and SLOs.

Skills

Go programming
Kubernetes
AWS
Systems thinking
Production systems ownership

Job description

Job Description

RedHat is seeking a Senior Software Engineer to join the HCP Platform Engineering team, building and operating ROSA (RedHat OpenShift Service on AWS) Hosted Control Planes (HCP). ROSA HCP is RedHat’s managed Kubernetes platform on AWS, built on a multi‑tenant architecture where RedHat operates shared control‑plane infrastructure while customers run workloads in their own AWS accounts. This is not a standard engineering role. You will write and ship production‑grade code, contribute to upstream open source projects and take ownership of production systems through on‑call. All three are equally important. You’ll work at the intersection of software engineering and production reliability on one of RedHat’s most complex and high‑scale platforms. The system spans multiple upstream open source projects and shared, multi‑tenant infrastructure, requiring strong engineering judgment and end‑to‑end ownership. The team is small, highly autonomous, and trusted to solve meaningful problems—from design through production. RedHat’s commitment to open source extends beyond our products into how we work. We foster a growth mindset and encourage engineers to thoughtfully and responsibly use AI to reduce toil, simplify workflows, and increase impact—allowing teams to focus on solving our customers’ hardest problems.

What you will do
  • Contribute production‑grade software to upstream open source projects including HyperShift and OpenShift, owning features end‑to‑end from design and implementation through deployment and long‑term lifecycle in production
  • Bring a product and systems lens to architecture decisions, ensuring designs account for scalability, operability, and real‑world production constraints from the start
  • Operate as a full‑stack systems owner, participating in on‑call rotations and taking end‑to‑end responsibility for diagnosing, fixing, and preventing production issues
  • Drive improvements that eliminate entire classes of failures by turning operational learning into durable product and platform enhancements
  • Design and evolve observability (metrics, logs, traces) and SLOs as part of the software lifecycle, ensuring systems are measurable, debuggable, and resilient by design
  • Raise the technical bar across the team through design docs, code review, pairing, and knowledge transfer during complex engineering work and incidents
  • Work in a high‑autonomy engineering team where you identify the most impactful problems and lead them from concept through implementation and production adoption
  • Partner as a peer with product and platform engineering teams to influence architecture, challenge assumptions, and ensure systems are built for scale, reliability, and long‑term operability
  • Integrate AI‑assisted development tools (GitHub Copilot, Cursor, Claude Code) into daily workflows for design, implementation, and debugging — using human judgment to maintain high engineering standards while increasing delivery velocity and system quality
What you will bring
  • 4+ years of software engineering experience building and shipping production systems in cloud environments, including microservices, platforms, or distributed systems
  • Strong proficiency in Go — you write production‑quality code, review it critically, and ramp quickly on large, unfamiliar codebases
  • Deep, hands‑on Kubernetes experience from a builder’s perspective: you’ve written operators, controllers, and CRDs in real‑world, multi‑tenant environments — not just operated clusters others built
  • Solid understanding of AWS fundamentals (EC2, IAM, networking) and how Kubernetes platforms behave and scale on AWS
  • Proven experience owning production systems under real SLOs, including participating in on‑call and leading incident response with a focus on root cause and long‑term fixes
  • You ramp fast on complex, unfamiliar systems — forming a mental model and making meaningful contributions within weeks
  • Highly self‑directed builder mindset: you identify high‑impact problems, propose solutions, and drive them end‑to‑end without waiting for direction
  • Strong systems thinking — you naturally connect design decisions to their downstream impact on scalability, reliability, and operability in production
  • Clear and effective communicator, able to collaborate with engineers on design, architecture, and tradeoffs
  • Nice to have: Experience with HyperShift, OpenShift, or ROSA in production environments
  • Familiarity with multi‑tenant Kubernetes challenges such as noisy neighbors, control plane scaling, and fleet‑level lifecycle management
  • Contributions to open source projects, particularly in the Kubernetes ecosystem
  • Experience designing and operating observability at scale (Prometheus, Grafana, Dynatrace, or similar)
  • Experience leveraging AI‑assisted development tools (e.g., coding agents, AI‑driven code review, spec‑driven workflows) to accelerate development and improve quality
Salary Range

$118,600.00 - $195,680.00. Actual offer will be based on your qualifications. Pay Transparency RedHat determines compensation based on several factors including but not limited to job location, experience, applicable skills and training, external market value, and internal pay equity. Annual salary is one component of RedHat’s compensation package. This position may also be eligible for bonus, commission, and/or equity. For positions with Remote‑US locations, the actual salary range for the position may differ based on location but will be commensurate with job duties and relevant work experience.

Benefits
  • Comprehensive medical, dental, and vision coverage
  • Flexible Spending Account – healthcare and dependent care
  • Health Savings Account – high deductible medical plan
  • Retirement 401(k) with employer match
  • Paid time off and holidays
  • Paid parental leave plans for all new parents
  • Leave benefits including disability, paid family medical leave, and paid military leave
  • Additional benefits including employee stock purchase plan, family planning reimbursement, tuition reimbursement, transportation expense account, employee assistance program, and more

Note: These benefits are only applicable to full time, permanent associates at RedHat located in the United States.

Equal Opportunity Policy (EEO)

RedHat is proud to be an equal opportunity workplace and an affirmative action employer. We review applications for employment without regard to their race, color, religion, sex, sexual orientation, gender identity, national origin, ancestry, citizenship, age, veteran status, genetic information, physical or mental disability, medical condition, marital status, or any other basis prohibited by law. RedHat does not seek or accept unsolicited resumes or CVs from recruitment agencies. We are not responsible for any fees, commissions, or payments related to unsolicited resumes or CVs, except as required in a written contract between RedHat and the recruitment agency or party requesting payment of a fee.

RedHat supports individuals with disabilities and provides reasonable accommodations to job applicants. If you need assistance completing our online job application, email application‑assistance@redhat.com. General inquiries, such as those regarding the status of a job application, will not receive a reply.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Red Hat, LLC • North Carolina

Hybrid
USD 118,600 - 195,680
Comprehensive medical, dental, and vis
Flexible Spending Account – healthcare
Health Savings Account
+5
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Red Hat, Inc. • Raleigh (NC)

Hybrid
USD 118,000 - 196,000
Medical, dental, vision
401(k) with match
Paid time off
+2
Technical Account Manager - Openshift
Technical Account Manager - Openshift

Red Hat, LLC • Minnesota

Hybrid
USD 96,000 - 154,000
Medical coverage
Dental & Vision
HSA/FSA
+5
Principal Software Engineer- Foreman Team
Principal Software Engineer- Foreman Team

Red Hat, LLC • Raleigh (NC)

On-site
USD 151,000 - 250,000
Medical coverage
Dental coverage
Vision coverage
+7
Principal Software Engineer - Red Hat OpenShift AI at Red Hat Raleigh, NC
Principal Software Engineer - Red Hat OpenShift AI at Red Hat Raleigh, NC

Red Hat • Raleigh (NC)

On-site
USD 148,000 - 246,000
Comprehensive benefits
401(k) with company match
Paid time off & parental leave
Principal Software Engineer- Foreman Team
Principal Software Engineer- Foreman Team

Red River • Raleigh (NC)

On-site
USD 152,000 - 250,000
Medical, dental, and vision coverage
401(k) with company match
Paid time off
+2
Architect, OpenShift - Active Top Secret Clearance
Architect, OpenShift - Active Top Secret Clearance

Red Hat Professional Consulting, Inc. (f.k.a Planning Technologies, Inc.) • Washington

Hybrid
USD 130,000 - 216,000
Medical, dental, vision
401(k) with employer match
Paid time off
+2
Security Technical Account Manager
Security Technical Account Manager

Red Hat, LLC • North Carolina

Hybrid
USD 107,000 - 173,000
Comprehensive medical, dental, and視on
Retirement 401(k) with employer match
Paid time off and holidays
+1
Architect, OpenShift - Active Top Secret Clearance
Architect, OpenShift - Active Top Secret Clearance

Red Hat • Washington

On-site
USD 130,000 - 215,000
Bonus eligibility
Paid time off
Retirement plan
+3
Senior OpenShift Consultant - Top Secret Clearance Required
Senior OpenShift Consultant - Top Secret Clearance Required

Red Hat Professional Consulting, Inc. (f.k.a Planning Technologies, Inc.) • Virginia (IL)

Hybrid
USD 130,000 - 215,000
Medical, dental, vision
401(k) matching
Paid time off
+1