On-Prem Reliability Engineer

OpsMill

Poland

Remote

PLN 180,000 - 300,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

OpsMill is seeking a Production Engineering / SRE-focused engineer to join our Infrahub team. You will work across on-prem deployments, diagnosing complex issues, building diagnostics, and advancing test automation to improve reliability.

You will collaborate with customers and internal teams, debugging Kubernetes environments, and contributing production-grade code in Python, Go, or Rust. This role emphasizes async communication, remote work capability, and cross-functional impact.

Qualifications

  • Bachelor's degree in a relevant field or equivalent experience.
  • Strong background in reliability engineering and production-grade software.
  • Experience with Kubernetes, troubleshooting distributed systems, and observability.

Responsibilities

  • Collaborate with customers and cross-functional teams on escalations across Kubernetes environments.
  • Reproduce issues, isolate root causes, and coordinate fixes with engineering.
  • Develop diagnostics tooling, health checks, and environment validators to speed up troubleshooting.
  • Own test automation roadmap, improve CI stability, and reduce flaky tests.
  • Establish performance baselines and regression tests to catch scale and latency issues early.
  • Improve installation and upgrade robustness through automation and guardrails.
  • Write production-quality code in Python, Go, or Rust for internal tooling.
  • Translate field issues into better tests, observability, and product defaults.

Skills

Production engineering
SRE practices
Kubernetes
Observability
Python/Go/Rust
Troubleshooting
Communication skills
Remote work capability

Education

Bachelor's degree in Computer Science

Tools

CI/CD tooling
Docker
Helm

Job description

OpsMill is seeking a Production Engineering / SRE-focused engineer to join our Infrahub team. You will work across on-prem deployments, diagnosing complex issues, building diagnostics, and advancing test automation to improve reliability.

You will collaborate with customers and internal teams, debugging Kubernetes environments, and contributing production-grade code in Python, Go, or Rust. This role emphasizes async communication, remote work capability, and cross-functional impact.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Balyasny Asset Management L.P. • Warszawa

On-site
PLN 180,000 - 300,000
Site Reliaibility engineer
Site Reliaibility engineer

Alcor • Kraków

On-site
PLN 250,000 - 380,000
Senior SRE: Kubernetes Production Reliability, Remote
Senior SRE: Kubernetes Production Reliability, Remote

Software Mind • Kraków

On-site
PLN 180,000 - 300,000
Flexible work
Remote work
Global projects
+6
Senior Site Reliability Engineer — Remote Incident Leader
Senior Site Reliability Engineer — Remote Incident Leader

Affirm • Poland

Remote
PLN 308,000 - 428,000
Parental benefits
Health care coverage
Flexible Spending Wallets
+2
Senior Platform SRE: Reliability & Observability Architect
Senior Platform SRE: Reliability & Observability Architect

IG KnowHow • Kraków

Hybrid
PLN 240,000 - 420,000
Growth opportunities
Mentoring programs
Networking clubs
+2
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Grid Dynamics • Województwo pomorskie

On-site
PLN 80,000 - 120,000
Medical insurance
Sports benefits
Professional development opportunities
+2
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Grid Dynamics • Kraków

On-site
PLN 254,000 - 340,000
Medical insurance
Sports benefits
Professional development opportunities
+2
Site Reliability Engineer
Site Reliability Engineer

Caspian One • Warszawa

On-site
PLN 180,000 - 280,000
Senior Platform Reliability Engineer
Senior Platform Reliability Engineer

IG Group • Kraków

Hybrid
PLN 250,000 - 420,000
Career development
Mentoring programs
Networking clubs
+3
SRE Engineer: Observability, AI-Driven Reliability
SRE Engineer: Observability, AI-Driven Reliability

Citibank (Switzerland) AG • Warszawa

Hybrid
Confidential
Pension plan
Private medical care
Life insurance
+1