Senior Site Reliability Engineer (m/w/d)

Impower

München

Vor Ort

EUR 70.000 - 90.000

Vollzeit

14 Tage+

Erhalte mehr Antworten von Arbeitgebern

Versende in nur wenigen Minuten einen passgenauen Lebenslauf.

Benefits dieser Stelle

Flexible hours
Ownership in projects
Diverse team culture

Zusammenfassung

Impower, located in Munich, is seeking a Senior Site Reliability Engineer to enhance the reliability of their AI-driven ERP platform. This role focuses on managing the platform's operational foundations across Kubernetes, AWS, CI/CD, and security.

The ideal candidate should have extensive experience in cloud environments, particularly with AWS and Kubernetes, and possess strong communication skills. Enjoy freedom and flexibility with a hybrid setup, allowing three days a week in the office.

Qualifikationen

  • 5+ years building and operating production systems in cloud environments.
  • Deep, hands-on Kubernetes experience and AWS knowledge.
  • Strong in Infrastructure as Code with Terraform.

Aufgaben

  • Own platform reliability end-to-end from AWS to Kubernetes.
  • Drive CI/CD excellence with GitLab and Terraform pipelines.
  • Manage scalable AWS infrastructure with strong IaC discipline.

Kenntnisse

Cloud Engineering
Kubernetes
AWS
Terraform
GitOps
Security Practices
Observability Tools
Scripting
Communication

Ausbildung

Bachelor's degree in Engineering or related field

Tools

Sentry
Grafana
Prometheus
Jenkins

Jobbeschreibung

Introduction

Munich | Hybrid | Product based company

At Impower, we are shaping the future ofproperty management — simple, fast, and digital. More than 100,000 people manage over 12 million apartments using processes that are still manual, complex, and time-consuming. We help property managers modernize these workflows step by step with reliable software, scalable systems, and real practical value.

As our platform continues to grow, reliability, scalability, security, and developer enablement become increasingly critical. We are building an AI-driven ERP platform that combines modern cloud infrastructure, workflow automation, and intelligent services to support complex operational processes at scale.

As aSenior Site Reliability Engineer, you will take ownership of the reliability and operational foundations of our platform. You will work across Kubernetes, AWS infrastructure, CI/CD, observability, and security to ensure that our systems remain scalable, resilient, and secure as we grow.

Your mission
  • Own platform reliability end-to-end:Co-own our Kubernetes-based platform on AWS alongside our current Senior SRE - ingress, autoscaling, service mesh, config and secrets — and make sure it scales as we grow.

  • Drive CI/CD excellence:Evolve our GitLab + Terraform + ArgoCD/Helm pipelines to deliver our Java/Spring Boot and React applications faster, safer, and with more self-service capability for product teams.

  • Manage cloud infrastructure:Design and operate scalable AWS infrastructure (EKS, RDS, ALB, IAM, VPC, S3) using Infrastructure as Code, with strong IaC discipline and clear change management.

  • Strengthen observability:Improve our Sentry, Grafana, Prometheus, and Loki setup so teams can define SLOs, debug fast, and operate their services with confidence.

  • Lead on security:Own our security posture across infrastructure and application layers — IAM, secrets management, network segmentation, container and dependency scanning, vulnerability management, supply chain security, and audit readiness. Embed security as a design constraint, not a bolted-on review step.

  • Improve incident response:Strengthen our on-call practices, runbooks, and post-incident learning. We treat reliability as a product feature.

  • Enable product teams:Provide tooling, guidance, and self-service capabilities that help product engineers adopt better operational and deployment practices — make the good path the easy path.

  • Support the broader platform surface:Temporal workflows, PostgreSQL operations, S3, our Estuary CDC pipeline, and AI service infrastructure on GCP/Azure as we expand our AI capabilities

Your profile
  • Engineering experience:5+ years building and operating production systems in cloud environments, including real ownership of non-trivial systems at scale.

  • Kubernetes depth:Deep, hands-on production Kubernetes experience — beyond kubectl apply, including operators, networking, autoscaling, and debugging.

  • AWS expertise:Strong working knowledge of EKS, RDS, ALB, IAM, VPC, S3, and the operational realities of running services on AWS.

  • Infrastructure as Code:Solid Terraform experience with disciplined IaC practices.

  • GitOps:Hands-on experience with ArgoCD, Helm, or equivalent declarative deployment tooling.

  • Security expertise applied to cloud-native environments:IAM best practices, secrets management, secure network architecture, container and dependency vulnerability scanning, secure SDLC principles, and familiarity with compliance frameworks (e.g. ISO 27001, SOC 2, or comparable). You proactively identify risks and contribute to incident response and audit readiness.

  • Observability instincts:You've built dashboards, defined SLOs, run real incidents, and used the resulting learning to improve systems.

  • Automation fluency:Comfortable scripting and building tooling in Python, Go, Bash, or similar.

  • Communication:Excellent written and verbal English (C1+). You document decisions, write runbooks people actually use, and explain tradeoffs clearly.

Nice to have
  • Experience with Temporal or other workflow orchestration systems

  • Exposure to CDC pipelines (Estuary, Debezium, or similar)

  • Multi-cloud experience (AWS primary, GCP/Azure for AI services)

  • Background running platforms for Spring/Java services at scale

  • Experience in regulated environments (financial services, real estate, healthcare)

  • Familiarity with AI/LLM infrastructure patterns or agentic engineering workflows

  • Prior experience as an early platform hire — building foundations without over-engineering

How you work
  • You think in systems and prefer building self-service capabilities over becoming a ticket queue

  • You're pragmatic about quality — protecting long-term adaptability without gold-plating

  • You communicate openly about tradeoffs, mistakes, and unknowns.

  • You see security and compliance as enablers of speed, not obstacles.

  • You're comfortable with autonomy and high ownership in a small, focused team.

  • You're curious about how AI is changing platform engineering — and you want to help figure it out.

Why us?

Freedom and flexibility:Hybrid setup (3 days / week in the office) from Munich, with flexible hours and real ownership.

Meaningful impact:Build technology that helps thousands of people simplify property management every day.

Modern environment:Work with a cutting-edge tech stack (React, TypeScript, Java, AWS, Kubernetes) and up-to-date tools.

Growth opportunities:Deep technical scope across SRE, cloud infrastructure, security, and automation — shaping the reliability and security foundations of the platform as we scale.

Supportive culture:Join a diverse, agile team that values autonomy, trust, and collaboration, with strong guidance during onboarding and beyond.

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.
oder ziehe deine Datei hierhin.
Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

(Senior) Site Reliability Engineer (m/f/d) in Berlin or Konstanz
(Senior) Site Reliability Engineer (m/f/d) in Berlin or Konstanz

United States Digital Space LLC • Deutschland

Hybrid
EUR 60.000 - 90.000
Hybrid working arrangements
Flexible hours
Subsidized sports or yoga courses
+1
Senior Backend Engineer (m/f/d) – Munich
Senior Backend Engineer (m/f/d) – Munich

Impower • München

Vor Ort
EUR 70.000 - 90.000
Flexible work hours
Ownership of projects
Growth opportunities
Senior Platform Engineer (AWS)
Senior Platform Engineer (AWS)

GotPhoto • Berlin

Hybrid
EUR 70.000 - 90.000
Unlimited paid holiday
Education budget
Remote work options
+1
Senior Site Reliability Engineer (all genders)
Senior Site Reliability Engineer (all genders)

FACT-Finder • Pforzheim

Hybrid
EUR 90.000 - 140.000
Hybrid work model
Impact on product reliability
Competitive compensation
+1
Senior Site Reliability Engineer (m/f/d)
Senior Site Reliability Engineer (m/f/d)

TOPdesk • Kaiserslautern

Vor Ort
EUR 90.000 - 150.000
30 days annual vacation
Remote-friendly options
Health and wellness programs
+2
Software Engineer
Software Engineer

edenity. • München

Hybrid
EUR 60.000 - 80.000
Direct impact on product
Access to cutting-edge AI tools
Flexibility and growth opportunities
Senior Site Reliability Engineer (x/f/m)
Senior Site Reliability Engineer (x/f/m)

United States Digital Space LLC • Berlin

Hybrid
EUR 90.000 - 140.000
Deutschlandticket
Vacation days
Health insurance
+7
Engineering Manager - Site Reliability & Observability (x/f/m)
Engineering Manager - Site Reliability & Observability (x/f/m)

United States Digital Space LLC • Berlin

Hybrid
EUR 120.000 - 180.000
Deutschlandticket (Germany-wide public
health insurance
Pension scheme (bAV)
+4
Senior Site Reliability Engineer (f/m/d)
Senior Site Reliability Engineer (f/m/d)

Personio • Berlin

Hybrid
EUR 90.000 - 130.000
Competitive reward package including 0
28 days of paid vacation +1 day after
Impact Day
+1
Senior Engineer (Platform)
Senior Engineer (Platform)

Jobgether • Deutschland

Vor Ort
EUR 120.000 - 160.000