Senior SRE: Bare-Metal Automation & Observability

CoreWeave

Warszawa

On-site

PLN 262,000 - 350,000

Full time

13 days ago
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Family-level Medical Insurance
Family-level Dental Insurance
Generous Pension Contribution
Life Assurance at 4x Salary
Critical Illness Cover
Employee Assistance Programme
Tuition Reimbursement
Work culture focused on innovative dis
ruption

Job summary

CoreWeave, The Essential Cloud for AI, is seeking a Senior Site Reliability Engineer for its MetalDev team to split focus 60/40 between production operations/reliability and engineering automation. You will lead incident responses, perform root-cause analyses, and contribute Go code, dashboards, and automated remediation workflows across global data centers.

Qualifications include 5+ years in SRE or related field, Go proficiency, Kubernetes, Prometheus and Grafana, and on-call experience.

Qualifications

  • Bachelor’s degree in Computer Science, Engineering, or a related technical field.
  • 5+ years of experience in Site Reliability Engineering, production engineering, cloud infrastructure, or software engineering.
  • Working proficiency in Go with experience developing production-quality software.
  • Hands-on production experience with Kubernetes and containerised microservices architectures.
  • Experience with observability and telemetry stacks, specifically Prometheus and Grafana.
  • Demonstrated track record supporting production services, leading incident management, and participating in on-call rotations.
  • Excellent troubleshooting, analytical, and technical documentation skills.

Responsibilities

  • Lead incident response, troubleshooting, root-cause analyses, and post-incident reviews.
  • Participate in an on-call rotation and ensure reliability across data centers.
  • Write resilient Go code and build Prometheus and Grafana dashboards.
  • Develop automated remediation workflows to reduce manual overhead across our fleet.
  • Define SLOs and KPIs, improve CI/CD deployment pipelines, and create self-service tooling for Fleet Operations and Hardware engineering teams.

Skills

Go
Kubernetes
Prometheus
Grafana
Incident management
On-call rotations

Education

Bachelor's degree in CS/Engineering

Tools

BMCs
Redfish

Job description

CoreWeave, The Essential Cloud for AI, is seeking a Senior Site Reliability Engineer for its MetalDev team to split focus 60/40 between production operations/reliability and engineering automation. You will lead incident responses, perform root-cause analyses, and contribute Go code, dashboards, and automated remediation workflows across global data centers.

Qualifications include 5+ years in SRE or related field, Go proficiency, Kubernetes, Prometheus and Grafana, and on-call experience.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer – Infra Automation
Site Reliability Engineer – Infra Automation

CoreWeave Europe • Warszawa

On-site
PLN 223,000 - 298,000
Family-level Medical Insurance
Family-level Dental Insurance
Generous Pension Contribution
+2
Senior SRE: Kubernetes Reliability for Global Platforms
Senior SRE: Kubernetes Reliability for Global Platforms

Software Mind • Kraków

Remote
PLN 180,000 - 240,000
Private healthcare and insurance
Multisport card
Language classes
+3
Senior SRE: Kubernetes Production Reliability, Remote
Senior SRE: Kubernetes Production Reliability, Remote

Software Mind • Kraków

On-site
PLN 180,000 - 300,000
Flexible work
Remote work
Global projects
+6
Senior Observability Engineer for AI Infrastructure
Senior Observability Engineer for AI Infrastructure

CoreWeave • Warszawa

On-site
PLN 321,000 - 428,000
Family-level Medical Insurance
Family-level Dental Insurance
Generous Pension Contribution
+5
Site Reliability Engineer
Site Reliability Engineer

Balyasny Asset Management L.P. • Warszawa

On-site
PLN 180,000 - 300,000
SRE Engineer: Observability, AI-Driven Reliability
SRE Engineer: Observability, AI-Driven Reliability

Citibank (Switzerland) AG • Warszawa

Hybrid
Confidential
Pension plan
Private medical care
Life insurance
+1
Senior Cloud SRE: Reliability, Observability & Automation
Senior Cloud SRE: Reliability, Observability & Automation

Renesas Electronics Corporation • Katowice

Hybrid
PLN 180,000 - 260,000
Remote Senior AI Hardware SRE - Automation & Reliability
Remote Senior AI Hardware SRE - Automation & Reliability

Akamai Career Site • Poland

Hybrid
PLN 180,000 - 320,000
FlexBase program
Senior Kubernetes SRE - Remote, High-Impact Reliability
Senior Kubernetes SRE - Remote, High-Impact Reliability

Engg • Poland

Remote
PLN 391,000 - 608,000
100% remote work
Flexible hours
International engineering team
+1
Senior Site Reliability Engineer, Infrastructure Engineering
Senior Site Reliability Engineer, Infrastructure Engineering

CoreWeave • Warszawa

On-site
PLN 262,000 - 350,000
Family-level Medical Insurance
Family-level Dental Insurance
Generous Pension Contribution
+6