(Senior) Site Reliability Engineer (m/f/d) in Berlin or Konstanz

KNIME

Berlin

Hybrid

EUR 90.000 - 130.000

Vollzeit

Vor 4 Tagen
Sei unter den ersten Bewerbenden
Bewerbungsgenerator

Verschicke keinen generischen Lebenslauf — erstelle einen Lebenslauf und ein Anschreiben, die genau auf diese Rolle zugeschnitten sind.

Schaffe es an den ATS-Filtern vorbei

Benefits dieser Stelle

Hybrid working
Flexible hours
Subsidised sports or yoga courses
Physiotherapy

Zusammenfassung

KNIME is building a next-generation cloud platform and seeks a Site Reliability Engineer to design, implement, and operate it at scale. You will own on-call rotations, drive reliability standards, and collaborate with multiple development teams to keep the SaaS stable, observable, and production-ready.

You will work on infrastructure as code, monitoring, and cost-efficient operations while interfacing with cloud providers like AWS and Azure.

Qualifikationen

  • Cloud certifications on AWS, Kubernetes, Linux or similar.
  • Strong cloud experience with AWS and/or Azure; mastery preferred.
  • Experience deploying software in Kubernetes environments; familiarity with patterns.
  • Scripting in Python or Shell required; Go/Java a plus.
  • Systems level knowledge of Linux; networking, security, load balancers, firewalls.
  • Telemetry: metrics, logging, tracing in large distributed systems.
  • Knowledge of OAuth/OIDC providers such as Keycloak.
  • Knowledge of PostgreSQL or similar relational databases.
  • Ability to work independently and with dispersed teams.

Aufgaben

  • Using code to automate deployment and operations of large-scale SaaS systems; Kubernetes operators a plus.
  • Build infrastructure as code with Helm, Terraform, CloudFormation, and Azure ARM.
  • Participate in on-call rotations, incident triage, troubleshooting and root-cause analysis.
  • Set standards for deployments including reliability, scalability, traceability and monitoring; collaborate with product teams.
  • Instrument deployed systems for performance, reliability and cost efficiency.
  • Embed with product and engineering to lead planning and drive adoption of reliability standards across teams.

Kenntnisse

Kubernetes
Terraform
AWS
Azure
Python
Shell
Go
Java
Linux
Networking
Monitoring
OAuth/OIDC

Tools

Kubernetes
Helm
Terraform
CloudFormation
Azure ARM
AWS
PostgreSQL
Keycloak

Jobbeschreibung

Mission
The Site Reliability Engineer at KNIME ensures that our next-generation cloud platform is built, operated, and scaled with reliability, security, and cost-efficiency as first-class engineering outcomes. Through active on-call ownership and hands-on engagement with engineering teams, this role sets and drives the operational standards that keep KNIME SaaS stable, observable, and production-ready at scale.

Mission
The Site Reliability Engineer at KNIME ensures that our next-generation cloud platform is built, operated, and scaled with reliability, security, and cost-efficiency as first-class engineering outcomes. Through active on-call ownership and hands-on engagement with engineering teams, this role sets and drives the operational standards that keep KNIME SaaS stable, observable, and production-ready at scale.

Mission
The Site Reliability Engineer at KNIME ensures that our next-generation cloud platform is built, operated, and scaled with reliability, security, and cost-efficiency as first-class engineering outcomes. Through active on-call ownership and hands-on engagement with engineering teams, this role sets and drives the operational standards that keep KNIME SaaS stable, observable, and production-ready at scale.
Role Overview
We are currently designing, building and launching our next generation of products and services in the cloud. We are looking for someone eager to be a part of this innovative process. This includes adapting our industry leading data science and analytics platform into a managed platform capable of serving thousands of users. The ability to handle Infrastructure as Code development is crucial as we strive to match our quality and functionality with innovative solutions that can address growth, cost and durability concerns. You will work in a cloud platform team interfacing with multiple development teams to drive KNIME products into production ready environments.

Responsibilities

  • Using code to automate the deployment and operations of large scale SaaS systems. Experience with Kubernetes operators is a plus.
  • Building out infrastructure as code using tools such as Helm, Terraform, Amazon CloudFormation and Azure ARM.
  • Participating in on-call rotations, incident triage and mitigation, troubleshooting issues in live environments, providing root cause analysis and issue resolution.
  • Setting standards for product deployments including reliability, scalability, traceability and monitoring. Communicate with product and development teams to help drive adoption.
  • Instrument deployed systems for performance, reliability and cost effectiveness.
  • Embeds with product and engineering to lead planning, own dependency risk, and drive consistent adoption of reliability and operational standards across teams.

Requirements

  • You hold one or more current certifications on a cloud platform such as AWS, Kubernetes, Linux or similar technologies
  • Strong cloud experience with at least one among AWS and Azure cloud providers. The ideal candidate would master both. You possess in-depth knowledge of VPC, IAM, EKS, ECR, EC2, S3, RDS, CloudWatch and their counterparts in the Azure environments.
  • Have experience deploying software systems to a Kubernetes environment. Have a working knowledge of Kubernetes concepts and the ability to craft deployment solutions using common Kubernetes patterns.
  • Scripting knowledge in Python, Shell are required, additional programming experience in Go, Java are a plus.
  • Systems level knowledge and experience with Linux. Expertise in networking, including security, routing, load balancers, and firewalls.
  • Knowledge of best practices around service telemetry, including metrics aggregation, distributed logging, and tracing in large, distributed systems
  • Working knowledge of OAuth/OIDC identity providers such as Keycloak
  • Working knowledge of relational databases such as Postgres
  • Ability to work independently and within a team environment. This includes clear and concise communication across an organization that is geographically and culturally dispersed

What Success Looks Like

  • Platform stability at scale: The multi-tenant SaaS platform maintains its availability SLO across all tenant tiers as the infrastructure grows from its current state toward a globally distributed commercial release.
  • Operational excellence embedded in engineering: Every team shipping to the platform follows a shared production readiness standard, reducing escaped defects and repeated incidents.
  • Automated, scalable operations: Tenant onboarding, deployments, and incident remediation are pipeline-driven, eliminating manual toil and enabling the team to scale without growing headcount at the same rate.
  • Commercial readiness: The platform meets enterprise security and compliance gates, supports consumption-based metering, and can sustain the onboarding velocity required for a commercial launch without operations becoming the bottleneck.

What We Offer

  • Purpose-driven impact: The opportunity to be a driving force and cloud advocate as we design and build our next generation of service offerings
  • Craft & collaboration: Work alongside experienced engineers in a systems-first culture that values simplicity, maintainability, and clean design.
  • Learning: Continuous growth through hands-on challenges, peer exchange, and exposure to cutting-edge AI and data analytics topics.
  • Flexibility, health and wellbeing: Hybrid working, flexible hours, subsidised sports or yoga courses, physiotherapy, and flu shots at select locations.

About Us
KNIME is a leading AI platform that enables organisations to make sense of their data through intuitive, scalable, and collaborative data science. We empower data professionals and business users alike to build, deploy, and manage AI and data workflows that drive better decisions. Hundreds of global enterprises use the KNIME platform including Citi, Bosch and P&G.
KNIME is an equal opportunity employer. We’re all about providing opportunities for different perspectives to come together, where everyone feels included no matter their background.

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.
oder ziehe deine Datei hierhin.
Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Senior Software Engineer (m/f/d) in Berlin or Konstanz
Senior Software Engineer (m/f/d) in Berlin or Konstanz

KNIME • Berlin

Hybrid
EUR 90.000 - 130.000
Hybrid working
Flexible hours
Wellbeing benefits
+1
Senior Software Engineer (m/f/d) in Berlin or Konstanz
Senior Software Engineer (m/f/d) in Berlin or Konstanz

KNIME AG • Berlin

Vor Ort
EUR 110.000 - 150.000
Hybrid working
Flexible hours
Subsidised sports or yoga courses
+2
Application Security Engineer (m/f/d) in Konstanz or Berlin
Application Security Engineer (m/f/d) in Konstanz or Berlin

KNIME • Berlin

Hybrid
EUR 90.000 - 120.000
Application Security Engineer (m/f/d) in Konstanz or Berlin
Application Security Engineer (m/f/d) in Konstanz or Berlin

KNIME AG • Berlin

Hybrid
EUR 65.000 - 85.000
Subsidized gym memberships
Flexible working hours
Continuous learning opportunities
Product Marketing Manager (Enterprise) (m/f/d) in Berlin, Konstanz or Remote Germany
Product Marketing Manager (Enterprise) (m/f/d) in Berlin, Konstanz or Remote Germany

KNIME • Berlin

Vor Ort
EUR 60.000 - 85.000
Senior HR Specialist - Recruiting & Onboarding (m/f/d) in Berlin or Konstanz
Senior HR Specialist - Recruiting & Onboarding (m/f/d) in Berlin or Konstanz

KNIME AG • Berlin, Konstanz

Vor Ort
EUR 60.000 - 75.000
Impact & Ownership
Modern HR Stack
Flexible Working Environment
+2
Senior HR Specialist - Recruiting & Onboarding (m/f/d) in Berlin or Konstanz
Senior HR Specialist - Recruiting & Onboarding (m/f/d) in Berlin or Konstanz

Meyandy LLC • Berlin

Hybrid
EUR 55.000 - 75.000
Hybrid working arrangements
AI-driven recruitment tools
Professional development opportunities
Senior HR Specialist - Recruiting & Onboarding (m/f/d) in Berlin or Konstanz
Senior HR Specialist - Recruiting & Onboarding (m/f/d) in Berlin or Konstanz

Jackalope Digital LLC • Berlin

Vor Ort
EUR 65.000 - 95.000
Hybrid working
Flexible hours
Global and inclusive culture
+1
Site Reliability Engineer Kubernetes (f/m/d)
Site Reliability Engineer Kubernetes (f/m/d)

Workwise • Karlsruhe

Hybrid
EUR 85.000 - 120.000
Hybrid-Arbeitsmodell
Flexible Arbeitszeiten
Kantine an einigen Standorten
+4
Site Reliability Engineer (m/f/d)
Site Reliability Engineer (m/f/d)

Solactive AG • Berlin

Vor Ort
Confidential
30 annual vacation days
Job ticket
Gym access
+2