Senior Reliability & Infrastructure Engineer - Remote

Jobgether

Lavamünd

Remote

EUR 117.000 - 261.000

Vollzeit

Vor 5 Tagen
Sei unter den ersten Bewerbenden
Bewerbungsgenerator

Eine vollständige Bewerbung in einer Minute — maßgeschneiderter Lebenslauf und Anschreiben, fertig zum Versenden.

Schaffe es an den ATS-Filtern vorbei

Benefits dieser Stelle

Equity participation
Fully remote work
Health, dental, and vision benefits
Flexible vacation policy
Open-source collaboration

Zusammenfassung

Jobgether is seeking a Senior Software Engineer - Reliability, Infrastructure, and Tooling to join a globally distributed, fully remote team. You’ll design, build, and operate scalable infrastructure for real-time and AI workloads, focusing on security, observability, and resilience.

You’ll work with product teams to ensure reliability, contribute to incident response, and advance internal tooling. Experience with Kubernetes, Linux networking, and open‑source infrastructure is valued; remote

Qualifikationen

  • Experience building and operating non-trivial production applications.
  • Significant experience with Kubernetes or an equivalent large-scale container orchestration platform.
  • Strong understanding of Linux internals and networking.
  • Proven experience with observability, monitoring, logging, and related tooling.
  • Experience operating large-scale, globally distributed systems.
  • Experience responding to and managing complex production incidents.
  • Experience operating open-source infrastructure technologies such as Kafka, ClickHouse, or comparable distributed systems.
  • Strong systems-thinking skills and an ability to reason about infrastructure in terms of signals, feedback, dependencies, and control mechanisms.
  • Strong communication and collaboration skills, particularly with partner engineering teams.
  • A pragmatic approach balancing immediate delivery with long-term maintainability.
  • A keen interest in observability, reliability engineering, automation, and reducing operational complexity.
  • Nice to have: data engineering and analytics.
  • Nice to have: global Layer 3 networking.
  • Nice to have: operating systems with long-lived workloads such as real-time media.
  • Nice to have: Google SRE or another high-scale reliability engineering environment.
  • Nice to have: PCI compliance experience.

Aufgaben

  • Ramp up on a complex global architecture involving distributed databases, messaging systems, networking infrastructure, Kubernetes, and other core platform technologies, identifying reliability debt and improvement.
  • Design and ship reliability-focused engineering work directly within production codebases, including load balancing, load shedding, instrumentation, scalability, and efficiency improvements.
  • Build and evolve internal infrastructure and developer tooling that enables product engineering teams to independently operate reliable workloads.
  • Partner closely with product development teams to co-design systems and ensure reliability, security, maintainability, and operational readiness are considered throughout development.
  • Develop observability capabilities that make system behavior measurable, understandable, and actionable, using appropriate signals and visualization techniques.
  • Participate in a shared on-call rotation and contribute to effective incident response, investigation, remediation, and prevention of recurring reliability issues.
  • Investigate complex system-level problems across distributed infrastructure, networking, application behavior, and production environments.
  • Improve configuration management and infrastructure practices across diverse systems, reducing unnecessary complexity, errors, and technical debt.
  • Contribute technical perspectives to architectural discussions and help establish engineering practices that support both short-term delivery and long-term scalability.
  • Support systems with demanding workloads, including real-time media, secure customer code execution, advanced networking, and other highly concurrent services.
  • Collaborate with engineering partners on potentially contentious reliability and operational decisions with clarity, pragmatism, and strong technical judgment.
  • Automate repetitive operational processes wherever possible to improve engineering efficiency and reduce manual intervention.

Kenntnisse

Kubernetes
Distributed systems
Observability
On-call experience
High concurrency
Security-minded
Reliability engineering
Open-source tooling
Real-time media

Tools

Kafka
ClickHouse
Prometheus
Grafana

Jobbeschreibung

Jobgether is seeking a Senior Software Engineer - Reliability, Infrastructure, and Tooling to join a globally distributed, fully remote team. You’ll design, build, and operate scalable infrastructure for real-time and AI workloads, focusing on security, observability, and resilience.

You’ll work with product teams to ensure reliability, contribute to incident response, and advance internal tooling. Experience with Kubernetes, Linux networking, and open‑source infrastructure is valued; remote

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.

oder ziehe deine Datei hierhin.

Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Senior Site Reliability Engineer — Remote, High-Impact
Senior Site Reliability Engineer — Remote, High-Impact

Jobgether • Österreich

Remote
EUR 90.000 - 150.000
Fully remote working environment
Global engineering team
Competitive compensation
Senior SRE - Kubernetes & Cloud Infra (Remote)
Senior SRE - Kubernetes & Cloud Infra (Remote)

Jobgether • Lavamünd

Vor Ort
EUR 148.000 - 201.000
100% remote work within EU time zones
Flexible working hours
High ownership and autonomy
+4
Senior Backend Engineer: Core APIs & Reliability
Senior Backend Engineer: Core APIs & Reliability

Far Coder • Lavamünd

Hybrid
EUR 70.000 - 150.000
Fully remote
Globally distributed team
Mentorship and growth
Remote Senior Cloud Infrastructure Engineer
Remote Senior Cloud Infrastructure Engineer

Jobgether • Österreich

Remote
EUR 90.000 - 130.000
Fully remote work
Senior Cloud Infrastructure Engineer
Senior Cloud Infrastructure Engineer

Jobgether • Lavamünd

Remote
EUR 148.000 - 211.000
Fully remote working environment
Broad technical ownership across cloud
Mentoring and leadership opportunities
+1
Senior SRE: Compute Nodes & Linux Infra
Senior SRE: Compute Nodes & Linux Infra

Jobgether • Lavamünd

Vor Ort
EUR 127.000 - 190.000
Competitive compensation
Career growth
Ownership of meaningful projects
+1
Remote Senior AI-First Full-Stack SaaS Engineer
Remote Senior AI-First Full-Stack SaaS Engineer

Jobgether • Lavamünd

Vor Ort
EUR 148.000 - 201.000
Fully remote
English classes
Premium medical insurance
+5
Senior Software Engineer - Reliability, Infrastructure, and Tooling
Senior Software Engineer - Reliability, Infrastructure, and Tooling

Jobgether • Lavamünd

Remote
EUR 117.000 - 261.000
Equity participation
Fully remote work
Health, dental, and vision benefits
+2
Senior Infrastructure Engineer
Senior Infrastructure Engineer

Jobgether • Österreich

Remote
EUR 90.000 - 130.000
Fully remote work
Senior Infrastructure Engineer
Senior Infrastructure Engineer

Jobgether • Lavamünd

Remote
EUR 148.000 - 211.000
Fully remote working environment
Broad technical ownership across cloud
Mentoring and leadership opportunities
+1