Senior Systems Software Engineer, Observability and Telemetry Platform

Jobtailor

Deutschland

Remote

EUR 90.000 - 130.000

Vollzeit

Vor 4 Tagen
Sei unter den ersten Bewerbenden
Bewerbungsgenerator

Eine zielgenaue Bewerbung für diesen Job — ein maßgeschneiderter Lebenslauf und ein Anschreiben, die genau zur Stellenanzeige passen.

Schaffe es an den ATS-Filtern vorbei

Zusammenfassung

Jobtailor is seeking an experienced engineer to design, implement, and support a large-scale Observability & Telemetry platform, focusing on performance at scale, real-time monitoring, logging, and alerting.

You will engage in the full lifecycle of services—from inception through deployment, operation, and refinement—while maintaining live systems, scaling via automation, and participating in on-call rotations with a strong emphasis on reliability and automation.

Qualifikationen

  • BS Degree in Computer Science (or related field).
  • 5+ years in infrastructure automation and distributed systems design.
  • Experience delivering observability platforms and tooling.

Aufgaben

  • Design, implement, and support the Observability & Telemetry platform at scale.
  • Engage in full service lifecycle from design to deployment and refinement.
  • Maintain live services by monitoring availability, latency and health.
  • Scale systems via automation to improve reliability and velocity.
  • Lead incident response and blameless postmortems; participate in on-call rotations.

Kenntnisse

Infrastructure Automation
Distributed Systems Design
Python Programming
Kubernetes Management
Observability Tools Experience

Ausbildung

BS Degree in Computer Science

Tools

Kubernetes
OpenStack
Docker
Grafana
OpenTelemetry
Prometheus

Jobbeschreibung

  • Design, implement, and support operational and reliability aspects of a large-scale Observability & Telemetry collection platform, focusing on performance at scale, real-time monitoring, logging, and alerting
  • Engage in and improve the full lifecycle of services, from inception and design through deployment, operation, and refinement
  • Support services before launch through system design consulting, development of software tools, platforms, and frameworks, capacity management, and launch reviews
  • Maintain live services by measuring and monitoring availability, latency, and overall system health
  • Scale systems sustainably through automation and evolve systems by driving changes that improve reliability and velocity
  • Practice sustainable incident response and blameless postmortems
  • Participate in an on-call rotation to support production systems
Requirements
  • BS degree in Computer Science or a related technical field involving coding (e.g., physics or mathematics), or equivalent experience
  • 5+ years of experience with Infrastructure automation, distributed systems design, and designing and developing tools for running large-scale private or public cloud systems in production
  • 5+ years of experience delivering foundational infrastructure and observability platforms
  • Experience with one or more of: Python, Go, Perl, or Ruby
  • In-depth knowledge of Linux, networking, and containers
  • Experience in using or running large private and public cloud systems based on Kubernetes, OpenStack, and Docker
  • Experience running Grafana, OpenTelemetry, Prometheus, and similar observability-focused tools
Core Competencies

Demonstrates expertise in designing and implementing large-scale Observability and Telemetry platforms, with a strong focus on performance, reliability, and automation. Proficient in managing cloud systems and utilizing observability tools to ensure system health and operational excellence.

Highest-signal resume keywords
  • Infrastructure Automation
  • Distributed Systems Design
  • Python Programming
  • Kubernetes Management
  • Observability Tools Experience
Hard Skills
  • Infrastructure Automation
  • Distributed Systems Design
  • Python
  • Go
  • Perl
  • Ruby
  • Linux
  • Networking
  • Containers
  • Cloud Systems
Soft Skills
  • Incident Response
  • Collaboration
Certifications & Qualifications
  • BS Degree in Computer Science
Industry Keywords
  • Observability
  • Telemetry
  • Performance Monitoring
  • System Health
  • Automation
Tools & Technologies
  • Kubernetes
  • OpenStack
  • Docker
  • Grafana
  • OpenTelemetry
  • Prometheus
Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.
oder ziehe deine Datei hierhin.
Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Senior Observability Engineer
Senior Observability Engineer

Jobtailor • Deutschland

Remote
EUR 90.000 - 120.000
Infrastructure Analyst – Page Diversity
Infrastructure Analyst – Page Diversity

Jobtailor • Deutschland

Remote
EUR 65.000 - 95.000
DevOps Engineer, Blockchain Infra
DevOps Engineer, Blockchain Infra

Jobtailor • Deutschland

Remote
EUR 70.000 - 110.000
Senior Engineer – Cloud
Senior Engineer – Cloud

Jobtailor • Deutschland

Remote
EUR 90.000 - 120.000
Senior Cloud Engineer – Infrastructure Services, FedRAMP
Senior Cloud Engineer – Infrastructure Services, FedRAMP

Jobtailor • Deutschland

Hybrid
EUR 110.000 - 160.000
Senior Infrastructure Engineer
Senior Infrastructure Engineer

Jobtailor • Deutschland

Remote
EUR 90.000 - 130.000
Lead OpenTelemetry Developer
Lead OpenTelemetry Developer

Embedded Shishya • Deutschland

Remote
EUR 90.000 - 120.000
SRE Monitoring Platform Software Engineer, Entry Level
SRE Monitoring Platform Software Engineer, Entry Level

Jobtailor • Deutschland

Remote
EUR 42.000 - 64.000
Mid-Level SRE Analyst
Mid-Level SRE Analyst

Jobtailor • Deutschland

Remote
EUR 90.000 - 130.000
Senior Software Engineer, DevOps
Senior Software Engineer, DevOps

Jobtailor • Deutschland

Remote
EUR 90.000 - 140.000