Staff Engineer (Core & MLOps)

Lever, Inc.

Belgique

À distance

EUR 120 000 - 180 000

Plein temps

Il y a 4 jours
Soyez parmi les premiers à postuler
Générateur de candidature

Obtenez une réponse de cet employeur — un CV et une lettre de motivation adaptés exactement à ce qu’on recherche pour ce poste.

Passez les filtres ATS

Avantages offerts par ce poste

Fully remote
Flexible hours
Global collaboration

Résumé du poste

Lever, Inc. is seeking a Staff Engineer (Core & MLOps) based in Belgium to shape foundational infrastructure for large-scale web data products and distributed teams. You will own architecture of core control and context planes enabling AI-driven workflows, across Kubernetes, Kafka, Java, Python, gRPC, and multi-cloud infra.

This is a fully remote, remote-first role with significant autonomy. You will mentor engineers, define service contracts, and drive reliability through SLOs, incident

Qualifications

  • 10+ years building scalable distributed backend systems.
  • Advanced Java with reactive frameworks (Vert.x/Netty) and strong Python.
  • Deep experience with gRPC and Protocol Buffers including schema evolution.
  • Hands-on production experience with Kubernetes, Terraform, and Kafka.
  • Experience designing telemetry pipelines or feedback systems from production data.
  • Strong reliability engineering background (SLOs/SLIs, fault tolerance).
  • Excellent technical writing and cross-team communication.
  • Experience in a globally distributed, remote-first environment.
  • MLOps experience including model serving or monitoring is a plus.

Responsabilités

  • Architect and evolve core control and context planes for multi-language services.
  • Maintain and improve Java/Python client libraries and deployment pipelines.
  • Define and govern inter-service contracts, including gRPC and API gateway policies.
  • Operate platform infrastructure across Kubernetes, Terraform, and data pipelines.
  • Lead architectural strategy via RFDs for orchestration and multi-cluster routing.
  • Establish reliability practices, SLOs, blast-radius analysis, and canary deployments.
  • Participate in on-call rotations and drive post-mortems for platform improvements.
  • Mentor engineers and enforce engineering standards across squads.

Connaissances

Advanced Java
Python proficiency
gRPC Protobuf
Kubernetes at scale
Terraform
Reliability engineering
Technical writing
MLops
Remote collaboration

Outils

Kubernetes
Terraform
Kafka
Envoy
Istio
SPIRE
Cilium
Helm

Description du poste

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Staff Engineer (Core & MLOps) based in Belgium.

This role offers the opportunity to shape foundational infrastructure powering large-scale web data products and distributed engineering teams. You’ll own the architecture of core control and context planes that enable services and AI-driven workflows to operate reliably and efficiently. Working across Kubernetes, Kafka, Java, Python, gRPC, and multi-cloud infrastructure, you’ll tackle complex distributed systems challenges at production scale. You’ll establish engineering standards, reliability practices, and service contracts that influence multiple product squads. The role combines hands-on architecture with technical leadership, mentoring, and cross-functional alignment. In a globally distributed, remote-first environment, you’ll have significant autonomy to solve challenging infrastructure problems and influence long-term platform strategy.

Accountabilities
  • Architect and evolve the control and context planes, advancing service and schema registries, SLO enforcement, health-aware routing, automated canary releases, and operational feedback loops.
  • Own the service chassis and golden path, maintaining and improving multi-language Java and Python client libraries, standardized workload specifications, Helm charts, and deployment pipelines.
  • Define and govern inter-service contracts, including gRPC and Protocol Buffer definitions, API gateway transcoding, versioning policies, and schema evolution standards.
  • Operate and improve the core platform infrastructure across Kubernetes, Terraform, HAProxy/Nginx, Confluent Kafka, real-time billing pipelines, Valkey, and database modernization initiatives.
  • Lead architectural strategy through Requests for Discussion (RFDs) covering workflow orchestration, gateway orchestration, multi-cluster routing, automated failover, and other critical platform initiatives.
  • Establish reliability engineering practices, including SLOs, SLIs, error budgets, fault isolation, and automated weighted canary deployments.
  • Participate in shared infrastructure on-call rotations, lead incident post-mortems, and convert operational insights into platform improvements.
  • Mentor engineers across multiple squads, review architectural proposals, and establish engineering practices that make reliable software development more consistent and efficient.
Requirements
  • 10+ years of experience building scalable distributed backend systems, with a strong track record of creating internal platforms or core libraries adopted across engineering organizations.
  • Advanced Java expertise, including reactive frameworks such as Vert.x or Netty, combined with strong Python proficiency.
  • Deep experience with gRPC and Protocol Buffers, including schema evolution and backward compatibility in mission-critical systems.
  • Hands-on production experience with Kubernetes at scale, Terraform, and event-streaming platforms such as Kafka.
  • Experience designing automated telemetry pipelines, materialized views, feature stores, or other feedback systems that use production data to dynamically improve system behavior.
  • Strong reliability engineering background, including SLO/SLI definition, blast-radius analysis, fault tolerance, and rigorous service contracts.
  • Exceptional technical writing skills and the ability to communicate complex architectural concepts clearly while driving alignment across teams.
  • Strong written and interpersonal communication skills suited to a globally distributed, remote-first environment.
  • A curious, continuous-learning mindset with an interest in evaluating new technologies, architectures, and engineering approaches.
  • Experience with Temporal, DBOS, or similar durable execution platforms is a plus.
  • MLOps experience, including model serving, performance monitoring, or production drift detection, is advantageous.
  • Familiarity with zero-trust networking and service meshes such as SPIRE, mTLS, Cilium, Istio, or Envoy is beneficial.
  • Experience building developer tooling such as CLIs, SDKs, or project generators is a plus.
  • Experience with large-scale web scraping or crawling, or contributions to distributed-systems and data-extraction open-source projects, is advantageous.
Benefits
  • Fully remote, remote-first working environment with flexible working hours.
  • Freedom and flexibility to work from the location where you are most productive.
  • Opportunity to work on core infrastructure supporting large-scale web data pipelines and distributed systems.
  • Exposure to cutting-edge open-source technologies, tools, and evolving AI and web data infrastructure.
  • Opportunities to attend conferences and connect with colleagues across the globe.
  • Collaboration with a diverse, multicultural, and globally distributed engineering community.
  • High level of autonomy and organizational trust.
  • Opportunities to influence platform architecture, engineering standards, and technical strategy across multiple teams.
Obtenez votre examen gratuit et confidentiel de votre CV.

ou faites glisser et déposez votre fichier ici.

Similar jobs

Postes similaires à comparer

Staff Engineer, Core & MLOps Platform (Remote)
Staff Engineer, Core & MLOps Platform (Remote)

Lever, Inc. • Belgique

À distance
EUR 120 000 - 180 000
Fully remote
Flexible hours
Global collaboration
DevOps Engineer
DevOps Engineer

Recurv • Brussel

Hybride
EUR 70 000 - 110 000
Brand-new office
Hybrid remote work
Cloud & MLOps Engineer - Financial Services
Cloud & MLOps Engineer - Financial Services

Ernst & Young Advisory Services Sdn Bhd • Diegem

Sur place
EUR 60 000 - 80 000
Extensive training on technical matters
Access to new technologies
Flexible working arrangements
Senior DevOps Engineer
Senior DevOps Engineer

Lever, Inc. • Belgique

À distance
EUR 70 000 - 110 000
Remote work flexibility
International projects
Visa/relocation support
+1
Lead Software Engineer
Lead Software Engineer

Swift Software • Terhulpen

Sur place
EUR 90 000 - 120 000
Competitive compensation
Learning & development programs
Internal mobility
+1
MLOps Engineer
MLOps Engineer

Faktion & Crius Group • Antwerpen

Sur place
EUR 60 000 - 80 000
Company car and fuel card or mobility budget
Comprehensive hospitalization and group insurance
Top-tier laptop and smartphone
+2
Scale-up Full-Stack DevOps Consultant
Scale-up Full-Stack DevOps Consultant

keystone-solutions • Brussel

Sur place
EUR 70 000 - 110 000
Consultancy Nature
Dynamic Projects
Turbo-Charged Learning
+2
Platform Engineer
Platform Engineer

Coreso SA • Brussel

Sur place
EUR 65 000 - 90 000
Mission with purpose
International team in Brussels
Hybrid working (40% office, 60% home)
+2
Platform Engineer
Platform Engineer

Coreso SA • Brussel Hoofdstad

Sur place
EUR 70 000 - 95 000
Hybrid work model
International team in Brussels
European energy impact
Senior Backend / Data Platform Engineer – Customer Data
Senior Backend / Data Platform Engineer – Customer Data

ING Belgium • Brussel

Sur place
EUR 70 000 - 95 000