Site Reliability Engineer - Cloud Operations

Swissquote Bank

Gland

Vor Ort

CHF 110.000 - 150.000

Vollzeit

vor 25 Stunden
Sei unter den ersten Bewerbenden
Bewerbungsgenerator

Mach aus dieser Rolle ein Vorstellungsgespräch — ein Lebenslauf und ein Anschreiben, die darauf ausgerichtet sind, was dieser Arbeitgeber sucht.

Schaffe es an den ATS-Filtern vorbei

Zusammenfassung

Swissquote Bank is seeking a Site Reliability Engineer to join the Cloud Operations team in Switzerland. You will migrate and modernize production apps on Kubernetes, integrate third-party software, and collaborate with software engineers to improve reliability and performance.

You will design and operate on our service mesh platform and implement canary deployments and progressive rollouts. You will define SLOs/SLIs, enhance observability across metrics/logs/traces, and explore AI tooling for

Qualifikationen

  • At least 3 years of experience in SRE, DevOps, Platform Engineering or similar production-focused role.

Aufgaben

  • Migrate and modernize production applications on Kubernetes.
  • Integrate third-party software into production platforms and ensure operational standards.
  • Collaborate with Software/IT Engineers to improve reliability and performance.
  • Design and operate applications on service mesh platforms.
  • Implement safe deployment patterns like canaries and progressive rollouts.
  • Define SLOs/SLIs and KPIs to drive improvements.
  • Improve observability with metrics/logs/traces.
  • Explore AI tools for troubleshooting and remediation.
  • Test system behavior under load and failures.
  • Automate repetitive ops tasks.
  • Provide Level-3 support and participate in on-call rotation.

Kenntnisse

Kubernetes
DevOps
SRE
CI/CD
Service mesh
Prometheus
Grafana
Elasticsearch/OpenTelemetry
Argo CD
Terraform

Tools

OpenShift
EKS
Istio
Linkerd
Envoy
Argo Rollouts
Argo Workflows
GCP/AWS/Azure

Jobbeschreibung

Site Reliability Engineer - Cloud Operations
  • Full-time
  • Managerial role: No

At Swissquote, we’re all in. All in to shake things up. All in to build the bank people actually want to use. All in to make finance less boring — and a lot more powerful.

We’re Switzerland’s leading digital bank — 1,400+ people across Europe, the Middle East and Asia, building real financial solutions for over a million clients worldwide. From trading and investing to everyday banking, we cover the full picture. We move fast, but we build things we’re genuinely proud of. Have a look behind the scenes by checkingHumans of Swissquote on Instagram.

Growing fast creates room. Room to try things, own things, and grow at a pace most places can’t offer. Whether you like the spotlight or prefer to just put your head down and do great work, there’s space for both here.

We’ve ditched the dress code — but never the chance to celebrate. Big win or small, we make it count. The kind of place where the atmosphere takes care of itself.

As an equal opportunity employer, we welcome candidates from all backgrounds, experiences and perspectives to join our team and contribute to our shared success.

Feeling it? It’s a good start.

Team Mission and Stakeholders

Are you passionate about Kubernetes, distributed systems and keeping production platforms reliable at scale?

Join our Cloud Operations team and help us migrate and modernize applications running at the core of Swissquote.

We’re looking for a Site Reliability Engineer who enjoys working close to production, solving reliability problems and improving how applications are deployed and operated.

In this role, you will:

  • Migrate and modernize production applications on Kubernetes,
  • Integrate third-party software into our production platforms and make it fit our operational standards,
  • Work alongside Software and IT Engineers to improve reliability, performance and operational readiness,
  • Design and operate applications on our service mesh platform,
  • Integrate safe deployment patterns such as canary releases and progressive rollouts,
  • Define SLOs, SLIs and useful operational KPIs, then use them to drive improvements,
  • Improve observability across metrics, logs and traces so problems are easier to spot and understand,
  • Explore and integrate AI tools that can help with troubleshooting, incident analysis and remediation,
  • Test how systems behave under load, during failures and when dependencies disappear,
  • Automate repetitive operational work whenever it makes sense,
  • Provide Level-3 support and participate in the on-call rotation.
  • At least 3 years of experience in SRE, DevOps, Platform Engineering or a similar production-focused role,
  • Solid hands-on experience running production workloads on Kubernetes, OpenShift, EKS or a similar Kubernetes platform,
  • Good knowledge of Helm and how to package, configure and maintain applications with it,
  • Experience working with service mesh technologies such as Istio or Linkerd, or strong Kubernetes networking experience,
  • A good understanding of service-to-service networking, traffic routing, mTLS and TLS,
  • Experience with GitOps and modern deployment strategies such as canary or progressive delivery,
  • A practical understanding of SRE concepts such as SLIs, SLOs and error budgets,
  • Experience with observability and tracing tooling such as Prometheus, Grafana, Elastic Stack or OpenTelemetry,
  • Strong Linux and networking fundamentals, including TCP/IP, DNS and load balancing,
  • Comfortable troubleshooting JVM-based applications in production and able to investigate issues related to heap usage, garbage collection or JVM configuration,
  • Comfortable automating things with Python, Go, Bash or another programming language,
  • Experience or strong interest in applying AI to observability, incident response or operational automation,
  • Experience or a strong interest in operating applications that depend on GPU resources or other AI infrastructure,
  • Familiarity with Infrastructure as Code tools such as Terraform, Ansible or Puppet.

Nice-to-Haves

  • Experience with Argo CD, Argo Rollouts or Argo Workflows,
  • Deeper experience with Istio, Linkerd or Envoy-based service mesh platforms,
  • Experience designing or operating Kubernetes platforms at scale,
  • Experience running Java or Spring Boot applications in production,
  • Hands-on experience tuning JVM applications for performance or low-latency workloads,
  • Experience integrating applications with self-hosted AI platforms such as vLLM,
  • Experience troubleshooting AI infrastructure integrations, including model access, GPU availability and NVIDIA MIG configurations.
  • Knowledge of Cilium, eBPF or other modern Kubernetes networking technologies,
  • Experience with public cloud or large private cloud environments,
  • A homelab, self-hosted services or side projects where you get to experiment, break things and build them again.

Who You Are

  • You like understanding why systems behave the way they do, especially when something goes wrong,
  • You automate repetitive work instead of accepting it as part of the job,
  • You’re comfortable working across development, infrastructure and operations teams,
  • You don’t mind getting deep into software you didn’t build yourself,
  • You’re curious about AI and where it can genuinely improve day-to-day operations,
  • You are fluent in English and have good conversational French,
  • You enjoy keeping up with cloud-native technologies and trying new approaches when they solve a real problem.

Please note that Swissquote never requests sensitive personal information or payment of any kind during the recruitment process. Any such request is fraudulent.

SQ2

By clicking the link above or any third-party link within this posting, you are leaving this site and going to a third-party website where the third-party website's terms and privacy policy apply

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.
oder ziehe deine Datei hierhin.
Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Site Reliability Engineer - eFX/ Crypto
Site Reliability Engineer - eFX/ Crypto

Swissquote Bank • Gland

Vor Ort
CHF 120.000 - 180.000
CKA/CKAD (Nice to Have)
Helm/GitHub Actions/ArgoCD
Kafka/RabbitMQ
Site Reliability Engineer: Cloud & Kubernetes at Scale
Site Reliability Engineer: Cloud & Kubernetes at Scale

Swissquote Bank • Gland

Vor Ort
CHF 110.000 - 150.000
Product Engineer
Product Engineer

Swissquote • Gland

Vor Ort
CHF 120.000 - 160.000
Site Reliability Engineer / Software Developer (open to all genders, 60–100%)
Site Reliability Engineer / Software Developer (open to all genders, 60–100%)

Intelliact AG • Zürich

Hybrid
CHF 80.000 - 110.000
5 weeks of vacation
Mobile plan valid throughout Europe
Free coffee & snacks
Senior System Engineer Kubernetes & Linux
Senior System Engineer Kubernetes & Linux

Jobup • Zürich

Vor Ort
CHF 140.000 - 200.000
Search & AI Visibility Specialist
Search & AI Visibility Specialist

Swissquote Bank • Gland

Vor Ort
CHF 140.000 - 200.000
Site Reliability Engineer (SRE) (w/m)
Site Reliability Engineer (SRE) (w/m)

Cassa Dei Medici • Genf

Vor Ort
CHF 120.000 - 180.000
Senior Java Full Stack Developer – DevOps & Cloud (AI-Augmented Engineering)
Senior Java Full Stack Developer – DevOps & Cloud (AI-Augmented Engineering)

Swiss Himmel GmbH • Basel

Vor Ort
CHF 180.000 - 210.000
Product Manager Associate - Trading
Product Manager Associate - Trading

Swissquote • Gland

Vor Ort
CHF 90.000 - 130.000
Senior Java Platform Engineer – Banking Sector
Senior Java Platform Engineer – Banking Sector

UNION BANCAIRE PRIVÉE, UBP SA • Genf

Vor Ort
CHF 180.000 - 260.000