Site Reliability Engineer, Full-Stack Performance

Rocket Homes Real Estate LLC

Seattle (WA)

On-site

USD 166,900 - 203,900

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Rocket Homes Real Estate LLC is seeking an embedded SRE to build self‑service infrastructure and guardrails for high‑volume feature teams. You will span Java/Spring services, Kubernetes, and the TypeScript/React layer to detect and fix reliability bottlenecks, and you will develop tooling to empower teams to own performance and reliability.

The role emphasizes full‑stack observability, AI‑driven diagnostics, and initiatives to modernize delivery pipelines, while collaborating with core teams to

Qualifications

  • 5+ years in SRE, Platform Engineering, or Backend/Full‑Stack with a reliability focus.
  • Proficiency in Java and TypeScript to navigate codebases and build internal tooling.
  • Experience with APM instrumentation, distributed tracing, or RUM.

Responsibilities

  • Full‑Stack Observability: build and instrument end‑to‑end observability stack including backend APM and frontend performance.
  • Operationalize AI: build plumbing for AI gateway including cost tracking and monitoring.
  • Intelligent Incident Response: integrate AI‑driven insights to automate diagnosis and reduce MTTR.
  • Resilient‑by‑Design Advisory: partner with teams to architect high‑traffic paths across stack.
  • Modernize Pipelines: optimize delivery pipelines to be observable and reliable for high‑traffic deployments.

Skills

Java/Spring
Kubernetes
TypeScript/React
Observability tooling
Distributed tracing
CI/CD pipelines
AI tooling
Datadog
AWS

Tools

Datadog
CI/CD tooling
AWS
CLI tools

Job description

About the Role

As an embedded SRE, you are a force multiplier. You won't be shipping product features or doing chores for other teams — you will be building the self‑service infrastructure and guardrails that allow Redfin's highest-volume feature teams to own their own speed and reliability. You operate across the full stack: from Java/Spring services and Kubernetes workloads to the TypeScript/React layer, you find where systems break under load and build the tooling so teams can see it themselves next time.

  • Full‑Stack Observability: Build and instrument the end‑to‑end observability stack — APM for backend services, RUM for frontend performance — so teams have a single, coherent picture of system health from request to render.
  • Operationalize AI: Build the "plumbing" for our AI gateway — cost tracking, usage monitoring, and reliability guardrails — so teams can leverage AI independently and safely.
  • Intelligent Incident Response: Integrate AI‑driven insights into our observability stack to automate diagnosis across the stack and help teams reduce MTTR.
  • Resilient‑by‑Design Advisory: Partner with core teams to architect high‑traffic paths that hold up under load — spanning service design, data access patterns, and delivery infrastructure.
  • Modernize Pipelines: Optimize delivery pipelines to be observable and reliable end‑to‑end, ensuring every deployment to high‑traffic environments is a non‑event.
About You

Operational Lead: You dive into the messy bugs and manual toil to understand them first, then build the tools and share the knowledge so those problems don't come back.

  • Full‑Stack Thinker: You're as comfortable tracing a slow database query through a Java service as you are profiling a JavaScript bundle. You find the real bottleneck, wherever it lives.
  • Systems Pragmatist: You understand how backend latency, CDN behavior, and frontend rendering interact to define the customer experience — and you design solutions that address the right layer.
  • Pragmatic Educator: You enjoy teaching teams to own their operational health through better tooling and standards.

5+ years in SRE, Platform Engineering, or Backend/Full‑Stack engineering with a strong reliability focus. Proficiency in Java and TypeScript — enough to navigate complex codebases across the stack, identify bottlenecks, and build internal tooling. Observability foundations: Experience with APM instrumentation, distributed tracing, or RUM (Core Web Vitals experience a plus, not a requirement). Automation mindset: Experience building CLI tools, automated workflows, or CI/CD pipelines. Bonus: Datadog, AWS/Kubernetes resource optimization, or managing third‑party AI service limits.

What you’ll get

Our team members fuel our strategy, innovation and growth, so we ensure the health and well‑being of not just you, but your family, too! We go above and beyond to give you the support you need on an individual level and offer all sorts of ways to help you live your best life. We are proud to offer eligible team members perks and health benefits that will help you have peace of mind. Simply put: We’ve got your back. Check out our full list of Benefits and Perks.

On‑Call Expectations

This role may include participation in an on‑call rotation to support production systems and ensure service reliability. On‑call responsibilities may include coverage during nights and weekends. If applicable, frequency and scheduling will be determined by team needs and communicated accordingly.

Compensation

The compensation for this position is $166,900.00-$203,900.00. The position may also be eligible for an annual bonus, incentives, and other employment‑related benefits including, but not limited to, medical, dental, and vision benefits, 401K retirement plan, and paid‑time off. More information regarding these benefits and others can be found here. The information regarding compensation and other benefits included in this paragraph is the company’s current, good faith estimate at the time of posting. [Compensation and benefits are subject to modification from time to time as the Company, in its sole and exclusive discretion, deems appropriate.] The Company may determine during its future reviews of the proposed compensation and benefits provided for this position that the compensation and benefits for such position should be reduced. In no event will the Company reduce the compensation for the position to a level below the applicable jurisdictional minimum wage rate for the position.

Location and Legal Disclaimers

Los Angeles County and San Francisco candidates only: qualified applicants with arrest or conviction records will be considered for employment per the Fair Chance Ordinance and the Fair Chance Initiative for Hiring.

Decisions related to employment are not based on race, color, religion, national origin, sex, physical or mental disability, sexual orientation, gender identity or expression, age, military or veteran status or any other characteristic protected by state and federal laws. The company provides reasonable accommodations to qualified individuals with disabilities in accordance with applicable state and federal laws. Applicants requiring reasonable accommodations in completing the application and/or participating in the application process should contact a member of the Human Resources team, at Careers@Rocket.com.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer, Full-Stack Performance
Site Reliability Engineer, Full-Stack Performance

Redfin • Seattle (WA)

On-site
USD 166,000 - 204,000
Medical, dental, and vision benefits
401K retirement plan
Paid time off
+1
Senior Software Developer - Sales Center Engineering
Senior Software Developer - Sales Center Engineering

Redfin • Seattle (WA)

On-site
USD 166,000 - 204,000
Health benefits
401K retirement plan
Paid-time off
Senior Software Developer, Platform Engineering (Hybrid or Remote)
Senior Software Developer, Platform Engineering (Hybrid or Remote)

Redfin • Seattle (WA)

On-site
USD 166,900 - 203,900
Software Engineer I
Software Engineer I

Rocket • Town of Texas (WI)

On-site
USD 84,000 - 207,000
Health benefits
401K retirement plan
Paid time off
Software Engineer II
Software Engineer II

Rocket • Michigan

On-site
USD 97,000 - 207,000
Software Developer I - Search
Software Developer I - Search

Rocket Homes Real Estate LLC • Seattle (WA)

On-site
USD 107,000 - 145,000
Medical benefits
Dental benefits
Vision benefits
+2
Staff Software Engineer
Staff Software Engineer

Rocket • Michigan

On-site
USD 149,000 - 318,000
Health benefits
401K plan
Paid time off
Staff Software Engineer
Staff Software Engineer

Worky • Michigan

On-site
USD 149,000 - 318,000
Health benefits
401K retirement plan
Paid time off
Senior Software Engineer I - Infrastructure
Senior Software Engineer I - Infrastructure

Rocket • Michigan

On-site
USD 143,000 - 239,000
Software Engineering Manager
Software Engineering Manager

Rocket Homes Real Estate LLC • Michigan

Hybrid
USD 123,000 - 276,000
Benefits package