Staff Engineer (Core & MLOps) - Remote

Zyte

United States

Remote

USD 180,000 - 240,000

Full time

8 days ago
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Zyte, a remote-first company, seeks a Staff Engineer to own the design and evolution of core platform architectures. The role focuses on building scalable backend systems and leading multi-cloud infrastructure across Core API, Browser, Edge, Antiban, AI, and Agentic products.

You will partner with the Team Lead, Senior Engineers, DevOps, and QA to influence release velocity and the architecture of several squads. This position offers global collaboration and impact across distributed teams.

Qualifications

  • 10+ years of experience building scalable distributed backend systems.

Responsibilities

  • Architect the control and context planes. Advance the control plane into a robust production substrate and design the context plane to aggregate operational signals for automated maintenance."
  • Own the service chassis and golden path. Maintain and refine multi-language client libraries, deployment pipelines, and standardized workload specifications.
  • Define inter-service contracts. Establish gRPC and Protocol Buffer definitions, API gateway transcoding, versioning policies, and schema evolution rules for safe deployments.
  • Operate the platform substrate. Collaborate with infra engineers to run Kubernetes across multiple clouds, Terraform, HAProxy/Nginx, Kafka, and ongoing DB migrations.
  • Drive architectural strategy via RFDs. Lead RFDs on durable workflow orchestration, per-request gateway orchestration, multi-cluster routing, and automated failover.
  • Institutionalize reliability engineering. Establish clear SLOs and error budgets, health-aware traffic isolation, and controlled canary deployments.
  • Share production ownership. Participate in on-call rotations, lead incident post-mortems, and translate lessons into platform improvements.
  • Elevate engineering standards. Mentor engineers across squads on distributed systems design and best practices.

Job description

Why Zyte?

At Zyte, we don’t just collect web data—we solve complex web data challenges at scale. We are a globally distributed team that is bold, curious, and dedicated to building innovative ways to deliver clean, reliable web data from across the internet.

  • Data is our passion: We unlock access to open web data, empowering organizations—from startups to global enterprises—to power competitive intelligence, analytics, and AI pipelines.
  • Remote-first culture: Work from anywhere in the world. With over 260 Zytans across 33 countries, our strength comes from diverse backgrounds, perspectives, and skills.
  • Engineering excellence: We love diving deep into complex code bases, evaluating emerging technologies, and pushing technical boundaries. If you thrive on creative problem-solving and reliably delivering production-grade solutions, you will fit right in.
  • Established system design framework: We systematically analyze functional and product requirements alongside architectural trade-offs to make informed, sustainable design decisions.
  • Real-world distributed systems challenges: Tackle rare engineering problems in custom networking, high-throughput streaming, and large-scale orchestration under real-world operational constraints.
  • Strong product ownership: Work on a profitable, market-leading platform with dedicated time and resources allocated for architectural research and strategic feature engineering.
Why This Role?

Zyte API powers a suite of Java and Python microservices running across multiple cloud providers and data centers. Every engineering team—spanning Core API, Browser, Edge, Antiban, AI, and Agentic products—relies on foundational infrastructure managed by our team: Kubernetes clusters, high-throughput Kafka pipelines, billing engines processing per-request metrics, and robust inter-service contracts.

We are evolving Zyte API into an automated software factory: a platform where developers and AI agents seamlessly create, validate, deploy, and maintain data extractors and workflows through governed interfaces. Two foundational layers drive this evolution. The control plane (service/schema registries, health-aware routing, automated canary releases) is actively under deployment. The context plane, which synthesizes observational signals to guide automated repair and optimizations, represents our next major design effort. As a Staff Engineer, you will own both architectures, setting the engineering standard for how services at Zyte are designed, built, and operated. You will be the technical lead in the Core & MLOps squad, collaborating closely with the Team Lead, Senior Engineers, DevOps, and QA, while directly influencing the release velocity and architecture of five neighboring squads.

Requirements
What You’ll Do
  • Architect the control and context planes. Advance the control plane (service registry, schema registry, SLO enforcement, CLI tooling) into a robust production substrate. Design and build the context plane to aggregate operational signals—such as domain extraction histories, IP reputation scores, and BigQuery cost/performance telemetry—creating automated feedback loops between intent, execution, and self-healing maintenance.
  • Own the service chassis and golden path. Maintain and refine multi-language client libraries (Java and Python), standardized workload specifications, Helm charts, and deployment pipelines, driving seamless adoption across all product teams.
  • Define inter-service contracts. Establish gRPC and Protocol Buffer definitions, API gateway transcoding, versioning policies, and schema evolution rules to enable autonomous, safe deployments across teams.
  • Operate the platform substrate. Partner with infrastructure engineers to run Kubernetes (across OCI, Hetzner, Servers.com, and GCP), Terraform, HAProxy/Nginx ingresses, Confluent Kafka, real-time event-billing pipelines, Valkey, and ongoing database modernizations (MySQL to PostgreSQL).
  • Drive architectural strategy via RFDs. Lead Requests for Discussion (RFDs) on critical platform initiatives, including durable workflow orchestration (Temporal/DBOS), per-request gateway orchestration, multi-cluster routing, and automated failover.
  • Institutionalize reliability engineering. Establish clear SLOs and error budgets for core capabilities, implementing health-aware traffic isolation and automated weighted canary deployments to eliminate manual release bottlenecks.
  • Share production ownership. Participate in the shared infrastructure on-call rotation, lead incident post-mortems, and translate operational lessons into platform improvements.
  • Elevate engineering standards. Mentor engineers across squads on distributed systems design, review system proposals, and establish intuitive best practices that make building reliable software effortless.
Who You Are
  • 10+ years of experience building scalable distributed backend systems, with a proven track record of authoring internal platforms or core …
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Staff Engineer - Core & MLOps for Distributed Systems
Staff Engineer - Core & MLOps for Distributed Systems

Zyte • United States

Remote
USD 180,000 - 240,000
Senior Product Growth Manager - Remote
Senior Product Growth Manager - Remote

Zyte • United States

Remote
USD 140,000 - 190,000
Senior DevSecOps Engineer
Senior DevSecOps Engineer

Zywave • United States

On-site
USD 90,000 - 120,000
Flexibility in work hours
Health and wellness programs
Professional development opportunities
Head of Forward Deployed Engineering
Head of Forward Deployed Engineering

Zep AI • San Francisco (CA)

On-site
USD 180,000 - 240,000
Engineering Manager, ZAIDYN Customer Engagement (ZCE) 2.0
Engineering Manager, ZAIDYN Customer Engagement (ZCE) 2.0

Zs Associates • New York (NY)

On-site
USD 170,000 - 250,000
Hybrid work model
Staff Engineer (Backend, DevOps, Infrastructure)
Staff Engineer (Backend, DevOps, Infrastructure)

Zuma • San Francisco (CA)

On-site
USD 130,000 - 160,000
Great health insurance, dental, and vision
Gym and workspace stipends
Unlimited PTO
Senior Director, AI Engineering
Senior Director, AI Engineering

ZS • South San Francisco (CA)

On-site
USD 230,000 - 360,000
Hybrid working model
Travel for client engagement
Staff Engineer (Core & MLOps)
Staff Engineer (Core & MLOps)

Jobgether SRL • United States

Remote
USD 140,000 - 190,000
Fully remote
Flexible working hours
Global collaboration
+1
Senior Architect
Senior Architect

Zizy Inc. • New York (NY), Northern (KY)

Hybrid
USD 110,000 - 140,000
Mental health support
Generous time off
Paid parental leave
+1
Full-Stack Software Engineer
Full-Stack Software Engineer

Zyphra • Palo Alto (CA)

On-site
USD 80,000 - 100,000
Comprehensive medical, dental, vision, and FSA plans
Competitive compensation and 401(k)
Relocation and immigration support
+2