Senior Site Reliability Engineer (SRE) – Kubernetes

Software Mind

Kraków

Remote

PLN 180,000 - 240,000

Full time

4 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Private healthcare and insurance
Multisport card
Language classes
Internal & external training
Well-being initiatives
Remote work flexibility

Job summary

Software Mind is seeking an experienced Site Reliability Engineer to join the AI Experience Framework team. You will own production reliability for a Kubernetes-based stack, participating in deployment, monitoring, and on-call activities.

You will collaborate with engineering teams to troubleshoot Node.js and JVM services, implement CI/CD, GitOps, and observability, and improve incident response, postmortems, and lead improvements.

Qualifications

  • 5+ years in SRE/DevOps or related role with Kubernetes experience
  • 3+ years Kubernetes production operations experience
  • Strong incident response experience including on-call and postmortems
  • Experience with Splunk for logs and troubleshooting
  • Proficient with Prometheus and Grafana dashboards and alerts
  • CI/CD and IaC for container deployments; Helm and GitOps tools
  • Solid Linux and networking fundamentals (DNS, load balancing, HTTP/2, Kubernetes networking)
  • Troubleshooting both Node.js and JVM/Java services
  • Experience with mTLS and JWT-based service authentication
  • Excellent written and spoken English

Responsibilities

  • Support deployment, operation, and reliability of production services on Kubernetes
  • Monitor service health and investigate production incidents
  • Participate in on-call support, RCAs, postmortems, and reliability improvements
  • Troubleshoot runtime, networking, and service-to-service issues with engineering teams
  • Support CI/CD, GitOps-based deployments, observability, and production monitoring
  • Work within client-directed backlog and priorities

Skills

Kubernetes operations
Incident response
CI/CD pipelines
GitOps
Monitoring & observability
Linux networking
Node.js and JVM
Security: mTLS/JWT
English communication

Tools

Kubernetes
Helm
ArgoCD
Flux
Splunk
Prometheus
Grafana
GitOps tooling
DNS/HTTP2 networking

Job description

Company Description

Software Mind develops solutions that make an impact for companies around the globe. Tech giants & unicorns, transformative projects, emerging technologies and limitless opportunities – these are a few words that describe an average day for us. Building cross-functional engineering teams that take ownership and crave more means we’re always on the lookout for talented people who bring passion and creativity to every project. Our culture embraces openness, acts with respect, shows grit & guts and combines employment with enjoyment.

Job Description
Project – the aim you'll have

We are the AI Experience Framework team that builds the platform powering ServiceNow's AI-first user interfaces - an SSR runtime (karuna) built on Lit and server-rendered web components, running behind a multi-tier proxy/HTTP2 routing chain with sharded V8 isolate pools, paired with a ServiceNow Glide/Java platform layer (karuna-glide) that supplies metadata, ACLs, and service artifacts. This role owns production reliability for that stack end to end: Kubernetes deployment and operations, observability, and hands-on troubleshooting of both the Node.js and JVM sides of the system - not generalist infrastructure work.

Position – How You’ll Contribute
  • Support the deployment, operation, and reliability of production services running on Kubernetes.
  • Monitor service health and investigate production incidents across distributed applications.
  • Participate in on-call support, incident response, root cause analysis, postmortems, and reliability improvements.
  • Troubleshoot application runtime, networking, and service-to-service issues in collaboration with engineering teams.
  • Support CI/CD, GitOps-based deployments, observability, and production monitoring.
  • Work within a client-directed backlog and established priorities.
Qualifications
Expectations – the experience you need
  • 5+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, Production Engineering, or a closely related role, including strong recent hands-on experience supporting Kubernetes-based production services.
  • 3+ years of hands-on production Kubernetes experience strongly preferred. Kubernetes production operations, including deployment, scaling, rollout / rollback, resource tuning, and service-to-service troubleshooting
  • Strong production incident response experience, including on-call, runbooks, postmortems, and paging hygiene
  • Splunk experience for log aggregation, search, and production troubleshooting
  • Prometheus and Grafana experience, specifically building alert rules and dashboards, not only using existing dashboards
  • CI/CD and infrastructure-as-code for containerized deployments, including Helm and GitOps tools such as ArgoCD or Flux
  • Strong Linux and networking fundamentals, including DNS, load balancing, TCP / HTTP, HTTP/2, and Kubernetes networking
  • Production troubleshooting experience across Node.js and JVM/Java services, with strong depth in at least one runtime environment. Experience may include Node.js heap snapshots, CPU profiling, event-loop and memory analysis, as well as JVM GC log analysis, thread dumps, JVM tuning, and Java service latency investigation.
  • Service-to-service authentication experience, including mTLS, certificate rotation, certificate format conversion, and JWT-based service authentication
  • Very good spoken and written English.
Additional Skills – The Edge You Have
  • Web Components / Lit experience, to perform first-level debugging of UI-related issues
  • Server-side rendering or isomorphic runtime experience
  • Canary rollout / multi-version production operations
  • Distributed tracing and request-context correlation
  • KEDA or event-driven autoscaling
  • Experience with enterprise platform integration layers
Our Offer – Professional Development, Personal Growth
  • Flexible employment and remote work
  • International projects with leading global clients
  • International business trips
  • Non-corporate atmosphere
  • Language classes
  • Internal & external training
  • Private healthcare and insurance
  • Multisport card
  • Well-being initiatives
Position at: Software Mind
#WORLD
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Grid Dynamics • Kraków

On-site
PLN 254,000 - 340,000
Medical insurance
Sports benefits
Professional development opportunities
+2
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Grid Dynamics • Województwo pomorskie

On-site
PLN 80,000 - 120,000
Medical insurance
Sports benefits
Professional development opportunities
+2
Senior Site Reliability Engineer - Remote
Senior Site Reliability Engineer - Remote

Akamai Technologies • Kraków

On-site
PLN 90,000 - 120,000
Health benefits
Financial benefits
Family support
+2
Senior SRE: Kubernetes Reliability for Global Platforms
Senior SRE: Kubernetes Reliability for Global Platforms

Software Mind • Kraków

Remote
PLN 180,000 - 240,000
Private healthcare and insurance
Multisport card
Language classes
+3
Site Reliability Engineer
Site Reliability Engineer

Balyasny Asset Management L.P. • Warszawa

On-site
PLN 180,000 - 300,000
Site Reliability Engineer
Site Reliability Engineer

Synechron • Kraków

On-site
PLN 180,000 - 300,000
Site Reliability Engineer
Site Reliability Engineer

Venquis • Poland

On-site
PLN 120,000 - 220,000
Senior SRE: Kubernetes Production Reliability, Remote
Senior SRE: Kubernetes Production Reliability, Remote

Software Mind • Kraków

On-site
PLN 180,000 - 300,000
Flexible work
Remote work
Global projects
+6
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Grid Dynamics • Wrocław

On-site
Competitive salary
Flexible schedule
Medical insurance
+4
Senior DevOps Engineer (Kubernetes & GitOps)
Senior DevOps Engineer (Kubernetes & GitOps)

Miratech • Województwo pomorskie

Remote
PLN 180,000 - 240,000
Health insurance
Relocation program
Work From Anywhere culture