Senior SRE Engineer (Observability Focus)

Capital

Warszawa

On-site

PLN 180,000 - 260,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Competitive Salary
Work-Life Harmony
Generous Time Off
Employee Referral Program
Health & Pension Benefits
Workation

Job summary

Capital in Warsaw is building out its observability practice and seeks a senior engineer who can own the end-to-end telemetry stack in a hybrid AWS and on‑premise environment. You will design, operate and optimize metrics, logs, and traces across Java, Python, and JS/TS services, mentor teammates, and drive reliability through dashboards, alerts and best practices.

This is a hands-on role that requires strong communication and a practical, data‑driven approach to incident response and

Qualifications

  • Experience owning full observability stack end-to-end.
  • Hands-on with metrics, logs, and traces in production.
  • Strong knowledge of VictoriaMetrics or Prometheus.
  • Experience with OpenSearch, ISM and ingest pipelines.
  • Proficiency in Kubernetes, GitOps, and Terraform/Ansible.
  • Proficient in Bash or Python for automation; English fluency.

Responsibilities

  • Own the full observability stack: metrics, logs, and traces from design to day-2 operations.
  • Architect VictoriaMetrics cluster topology and related scraping/remote write setups.
  • Operate OpenSearch clusters with ISM, hot-warm-cold, and ingest pipelines.
  • Build OTEL Collector pipelines and instrument services across languages.
  • Run Kafka as the telemetry transport and optimize throughput.
  • Manage log shipping with Fluent Bit/Vector/Fluentd and standardize structured logging.
  • Create Grafana dashboards and alerts that engineers actually use.
  • Collaborate to improve sampling, batching, and context propagation.
  • Contribute to incident response, post-mortems, and reliability improvements.
  • Mentor engineers on observability practices and tooling.

Skills

VictoriaMetrics
OpenSearch
OpenTelemetry
Kafka
Fluent Bit
Vector
Fluentd
Kubernetes
Argo CD
Terraform
Ansible
Bash
Python
Java
JavaScript
English

Tools

VictoriaMetrics
OpenSearch
Data Prepper
OTEL Collector
Grafana
PromQL/SQL
Prometheus
Kafka
Grafana Loki

Job description

Overview

We are a leading trading platform that is ambitiously expanding to the four corners of the globe. Our top-rated products have won prestigious industry awards for their cutting‑edge technology and seamless client experience. We deliver only the best, so we are always in search of the best people to join our ever‑growing talented team.

We’re building out our observability practice and need a senior engineer who can own it end to end. This is a hands‑on role. You’ll design and operate the telemetry stack that gives our engineering teams real visibility into production — across a hybrid AWS and on‑premise environment, at scale.

Responsibilities & Qualifications
  • Own the full observability stack: metrics (VictoriaMetrics), logs (OpenSearch), and traces (OpenTelemetry) — from pipeline design to day‑2 operations.
  • Architect and run VictoriaMetrics cluster topology (vmstorage/vminsert/vmselect), including vmagent scraping, remote write configuration, vmalert rules, and cardinality control.
  • Operate OpenSearch clusters: index lifecycle management (ISM), hot‑warm‑cold architecture, shard tuning, and ingest pipelines via Data Prepper.
  • Build and maintain OTEL Collector pipelines — receivers, processors, exporters — and instrument services across Java, Python, and JS/TS stacks (auto and manual).
  • Run Kafka as the telemetry transport layer (OTEL Collector → Kafka → backends), including topic design, partition strategy, consumer group lag monitoring, and throughput tuning for high‑volume telemetry.
  • Manage log shipping infrastructure using Fluent Bit, Vector, or Fluentd; define structured logging standards and field normalization across services.
  • Build Grafana dashboards and alerting that engineers actually use — clear, actionable, with well‑structured variables and thresholds.
  • Work with platform and application teams to improve sampling strategies (head/tail), batching, and context propagation across distributed services.
  • Contribute to incident response, post‑mortems, and reliability improvements driven by observability signals.
  • Mentor engineers on observability practices, tooling, and structured logging standards.
  • 6+ years in a DevOps, SRE, or platform engineering role, with at least 2 years focused on observability tooling at production scale.
  • Deep hands‑on experience with VictoriaMetrics (or Prometheus) — MetricsQL/PromQL, exporters, service discovery, remote write, downsampling, and retention management.
  • Solid OpenSearch or Elasticsearch skills: cluster operations, Query DSL, ISM policies, and ingest pipeline design.
  • Production experience with OpenTelemetry: Collector configuration, OTLP, context propagation, and instrumentation across multiple languages.
  • Strong Kafka skills — producer/consumer patterns, consumer group management, Kafka Connect, Schema Registry, and JMX‑based monitoring. Strimzi experience a plus if you’ve run Kafka on Kubernetes.
  • Proficiency with log shippers (Fluent Bit, Vector, Fluentd) and structured log parsing/normalization.
  • Working knowledge of Kubernetes (operators, Helm), Argo CD/GitOps, and Terraform/Ansible.
  • Comfortable in a hybrid AWS + on‑prem environment; solid understanding of networking as it applies to scraping and shipping pipelines.
  • Scripting ability in Bash or Python for automation and tooling.
  • Strong communication skills — you can explain observability tradeoffs clearly to engineers and non‑engineers alike.
  • English proficiency.
Benefits
  • Competitive Salary: We believe great work deserves great pay! Your skills and talents will be rewarded with a salary that makes you feel valued and motivated.
  • Work-Life Harmony: Join a company that genuinely cares about you — because your life outside of work matters just as much as your time on the clock. #LI-Hybrid
  • Generous Time Off: Need a breather? Our annual leave policy lets you recharge and enjoy life outside of work without a worry.
  • Employee Referral Program: Love working here? Share the love! Bring your talented friends on board and get rewarded for growing our awesome team.
  • Comprehensive Health & Pension Benefits: From medical insurance to pension plans, we’ve got your back. Plus, location‑specific benefits and perks!
  • Workation Wonderland: Live your digital nomad dreams with 30 extra days to work remotely from anywhere in the world (some restrictions apply). Adventure awaits!
  • Volunteer Days: Make a difference! Take two additional paid days each year to support causes you care about and give back to the community.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

DevOps Engineer (SRE Observability)
DevOps Engineer (SRE Observability)

Lingaro • Poland

Hybrid
PLN 180,000 - 260,000
Stable employment
Office as an option
Workation
+2
DevOps Engineer
DevOps Engineer

Duco Technology Ltd • Wrocław

On-site
PLN 180,000 - 240,000
Competitive salary
Commission bonus
Private Medical Insurance - Enel Med
+7
Lead DevOps Engineer — Observability & CI/CD Champion
Lead DevOps Engineer — Observability & CI/CD Champion

Duco Technology Ltd • Wrocław

On-site
PLN 180,000 - 240,000
Senior DevOps Engineer / Platform Engineer - Streaming, Caching, DBaaS
Senior DevOps Engineer / Platform Engineer - Streaming, Caching, DBaaS

SentinelOne • Poland

On-site
PLN 200,000 - 320,000
Restricted Stock Units (RSUs)
Employee Stock Purchase Plan (ESPP)
Competitive leave benefits
+7
Senior Platform SRE
Senior Platform SRE

IG Group • Kraków

Hybrid
PLN 240,000 - 360,000
Tailored development programs
Mentoring by leaders
Career progression
Site Reliability Engineer
Site Reliability Engineer

Balyasny Asset Management L.P. • Warszawa

On-site
PLN 180,000 - 300,000
Senior Staff Software Engineer
Senior Staff Software Engineer

Kontakt.io • Kraków

On-site
PLN 212,404 - 297,366
Equity in Series C company
Private healthcare
Multisport card
+1
Senior Observability SRE Engineer — Remote/Hybrid
Senior Observability SRE Engineer — Remote/Hybrid

Capital • Warszawa

Hybrid
PLN 180,000 - 260,000
Competitive Salary
Work-Life Harmony
Generous Time Off
+3
Senior Staff Software Engineer
Senior Staff Software Engineer

Kontakt Micro-Location Sp. Z.o.o. • Kraków

Hybrid
PLN 180,000 - 260,000
Equity
Private healthcare
Hybrid from Kraków office
+2
Senior Site Reliability Engineer (SRE) – Kubernetes
Senior Site Reliability Engineer (SRE) – Kubernetes

Software Mind • Kraków

Remote
PLN 180,000 - 240,000
Private healthcare and insurance
Multisport card
Language classes
+3