Staff Engineer (Core & MLOps)

Lever, Inc.

Schweiz

Remote

CHF 140.000 - 190.000

Vollzeit

Vor 2 Tagen
Sei unter den ersten Bewerbenden
Bewerbungsgenerator

Verschicke keinen 08/15-Lebenslauf — erstelle einen Lebenslauf und ein Anschreiben, die genau auf diese Rolle zugeschnitten sind.

Schaffe es an den ATS-Filtern vorbei

Benefits dieser Stelle

Fully remote
Flexible hours
Work on core infra for large-scale web
Open source involvement
Conference presence
Global collaboration
High autonomy
Influence platform architecture

Zusammenfassung

Lever, Inc. is seeking a Staff Engineer (Core & MLOps) based in Switzerland to shape foundational infrastructure powering large-scale web data products and distributed engineering teams.

You will own the architecture of core control planes for services and AI-driven workflows to operate reliably and efficiently, working across Kubernetes, Kafka, Java, Python, gRPC, and multi-cloud stacks with significant autonomy.

Qualifikationen

  • 10+ years building scalable distributed backend systems.
  • Advanced Java with reactive frameworks and strong Python.
  • Deep experience with gRPC and Protocol Buffers.
  • Production Kubernetes at scale with Kafka.
  • Reliability engineering and service contracts.

Aufgaben

  • Architect and evolve control and context planes, enforcement of SLOs and canary releases.
  • Own the service chassis, Java/Python client libraries, standardized workload specs, deployments.
  • Define inter-service contracts, including gRPC/Protobuf, API gateway transcoding, versioning.
  • Operate core platform infra across Kubernetes, Terraform, Nginx/HAProxy, Confluent Kafka, real-time pipelines.
  • Lead architectural strategy via RFDs for workflow orchestration and routing.
  • Establish reliability practices: SLOs/SLIs, fault isolation, automated canaries.
  • Mentor engineers, review proposals, promote reliable software practices.
  • Participate in on-call rotations and post-mortems to improve platforms.

Kenntnisse

Advanced Java
Python
gRPC/Protobuf
Kubernetes
Terraform
SRE practices
Technical writing
Cross-team collaboration

Tools

Helm
Kafka
Nginx/HAProxy
Confluent Kafka
Protobuf
SPIRE/Envoy

Jobbeschreibung

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Staff Engineer (Core & MLOps) based in Switzerland.

This role offers the opportunity to shape foundational infrastructure powering large-scale web data products and distributed engineering teams. You’ll own the architecture of core control and context planes that enable services and AI-driven workflows to operate reliably and efficiently. Working across Kubernetes, Kafka, Java, Python, gRPC, and multi-cloud infrastructure, you’ll tackle complex distributed systems challenges at production scale. You’ll establish engineering standards, reliability practices, and service contracts that influence multiple product squads. The role combines hands‑on architecture with technical leadership, mentoring, and cross‑functional alignment. In a globally distributed, remote‑first environment, you’ll have significant autonomy to solve challenging infrastructure problems and influence long‑term platform strategy.

Accountabilities
  • Architect and evolve the control and context planes, advancing service and schema registries, SLO enforcement, health‑aware routing, automated canary releases, and operational feedback loops.
  • Own the service chassis and golden path, maintaining and improving multi‑language Java and Python client libraries, standardized workload specifications, Helm charts, and deployment pipelines.
  • Define and govern inter‑service contracts, including gRPC and Protocol Buffer definitions, API gateway transcoding, versioning policies, and schema evolution standards.
  • Operate and improve the core platform infrastructure across Kubernetes, Terraform, HAProxy/Nginx, Confluent Kafka, real‑time billing pipelines, Valkey, and database modernization initiatives.
  • Lead architectural strategy through Requests for Discussion (RFDs) covering workflow orchestration, gateway orchestration, multi‑cluster routing, automated failover, and other critical platform initiatives.
  • Establish reliability engineering practices, including SLOs, SLIs, error budgets, fault isolation, and automated weighted canary deployments.
  • Participate in shared infrastructure on‑call rotations, lead incident post‑mortems, and convert operational insights into platform improvements.
  • Mentor engineers across multiple squads, review architectural proposals, and establish engineering practices that make reliable software development more consistent and efficient.
Requirements
  • 10+ years of experience building scalable distributed backend systems, with a strong track record of creating internal platforms or core libraries adopted across engineering organizations.
  • Advanced Java expertise, including reactive frameworks such as Vert.x or Netty, combined with strong Python proficiency.
  • Deep experience with gRPC and Protocol Buffers, including schema evolution and backward compatibility in mission‑critical systems.
  • Hands‑on production experience with Kubernetes at scale, Terraform, and event‑streaming platforms such as Kafka.
  • Experience designing automated telemetry pipelines, materialized views, feature stores, or other feedback systems that use production data to dynamically improve system behavior.
  • Strong reliability engineering background, including SLO/SLI definition, blast‑radius analysis, fault tolerance, and rigorous service contracts.
  • Exceptional technical writing skills and the ability to communicate complex architectural concepts clearly while driving alignment across teams.
  • Strong written and interpersonal communication skills suited to a globally distributed, remote‑first environment.
  • A curious, continuous‑learning mindset with an interest in evaluating new technologies, architectures, and engineering approaches.
  • Experience with Temporal, DBOS, or similar durable execution platforms is a plus.
  • MLOps experience, including model serving, performance monitoring, or production drift detection, is advantageous.
  • Familiarity with zero‑trust networking and service meshes such as SPIRE, mTLS, Cilium, Istio, or Envoy is beneficial.
  • Experience building developer tooling such as CLIs, SDKs, or project generators is a plus.
  • Experience with large‑scale web scraping or crawling, or contributions to distributed‑systems and data‑extraction open‑source projects, is advantageous.
Benefits
  • Fully remote, remote‑first working environment with flexible working hours.
  • Freedom and flexibility to work from the location where you are most productive.
  • Opportunity to work on core infrastructure supporting large‑scale web data pipelines and distributed systems.
  • Exposure to cutting‑edge open‑source technologies, tools, and evolving AI and web data infrastructure.
  • Opportunities to attend conferences and connect with colleagues across the globe.
  • Collaboration with a diverse, multicultural, and globally distributed engineering community.
  • High level of autonomy and organizational trust.
  • Opportunities to influence platform architecture, engineering standards, and technical strategy across multiple teams.
Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.

oder ziehe deine Datei hierhin.

Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Staff Engineer: Core Infra & MLOps (Remote)
Staff Engineer: Core Infra & MLOps (Remote)

Lever, Inc. • Schweiz

Remote
CHF 140.000 - 190.000
Fully remote
Flexible hours
Work on core infra for large-scale web
+5
Engineering Manager - Foundations & Enablement
Engineering Manager - Foundations & Enablement

Lever, Inc. • Schweiz

Remote
CHF 170.000 - 250.000
Remote stipend
Home office package
Office upgrades
+7
Senior Backend Engineer: Machine Learning Infrastructure
Senior Backend Engineer: Machine Learning Infrastructure

Lever, Inc. • Schweiz

Remote
CHF 67.000 - 100.000
Fully remote
Unlimited vacation
Home-office stipend
+6
Senior Full Stack Engineer - Core UX
Senior Full Stack Engineer - Core UX

Lever, Inc. • Schweiz

Remote
EUR 71.000 - 145.000
Equity options
Paid time off
Global team offsites
+4
Software Engineer – Large-Scale Data Processing Platforms
Software Engineer – Large-Scale Data Processing Platforms

Albedis SA • Bern

Vor Ort
CHF 110.000 - 140.000
Partial remote work
Modern workspace
International environment
+6
DevOps Engineer
DevOps Engineer

AlpineAI AG • Zürich

Hybrid
CHF 140.000 - 190.000
Shares program
Own hardware access
Senior Data Engineer (Python / AWS / ML Pipelines)
Senior Data Engineer (Python / AWS / ML Pipelines)

Lever, Inc. • Schweiz

Vor Ort
CHF 120.000 - 180.000
Collegial culture
Agile environment
Learning opportunities
+4
Senior Software Engineer, Quality
Senior Software Engineer, Quality

Lever, Inc. • Schweiz

Remote
CHF 120.000 - 194.000
Remote work
Home office budget
VSOP equity
+1
Automation Engineer (Customer Service)
Automation Engineer (Customer Service)

Lever, Inc. • Schweiz

Vor Ort
CHF 110.000 - 150.000
Competitive compensation
International teams
Flexible working environment
+1
Senior System Engineer (Cloud Native, 80-100%, Remote-First) - ID260507
Senior System Engineer (Cloud Native, 80-100%, Remote-First) - ID260507

Adfinis • Bern

Hybrid
CHF 120.000 - 170.000
Open Source culture
Flexible work arrangements
Tool freedom (Emacs/Vi etc.)
+2