Remote ML Data Infrastructure Engineer

Bayside Solutions

Cupertino (CA)

Remote

USD 83,000 - 96,000

Full time

12 days ago
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Bayside Solutions, Inc. is seeking a Software Engineer to build and operate the data infrastructure feeding ML training and inference systems. You will design high-throughput, distributed data pipelines atop columnar formats to ensure data is readily available for GPUs and models.

The role emphasizes hands-on performance engineering, collaboration with ML researchers, and designing reliable, observable data infrastructure with cost awareness. W2 contract, remote-friendly in Cupertino, CA.

Qualifications

  • Extensive experience building and running large-scale distributed data or ML infrastructure in production.
  • Strong programming skills in Python, plus a systems language for performance-critical work (Rust strongly preferred; C++ or Go acceptable).
  • Deep familiarity with columnar and lakehouse formats (Parquet, Iceberg, Delta, or Lance) and the trade-offs between them.
  • Hands-on performance engineering for I/O-bound workloads: Arrow, zero-copy, memory mapping, async I/O, and high-throughput object storage access patterns.
  • Working knowledge of the end-to-end ML workflow and how training and inference workloads consume data, enough to design data systems that serve them well.

Responsibilities

  • Design, build, and operate large-scale distributed data systems that serve ML training and inference workloads in production.
  • Strong Python Engineer.
  • in Rust (or C++/Go) with Python bindings for ML practitioners to build high-performance data loading, storage, and retrieval layers (Systems Performance, which is critical work)
  • Evaluate and adopt columnar and lakehouse formats (Parquet, Iceberg, Delta, Lance), and make the trade-offs explicit for schema evolution, random access, scan performance, and versioning.
  • Optimize I/O-bound pipelines using Arrow, zero-copy techniques, memory mapping, async I/O, and efficient object storage access patterns (request coalescing, prefetching, caching, parallel range reads).
  • Profile and remove bottlenecks across the data path, from object storage to host memory to accelerator, so training and inference stay compute-bound rather than I/O-bound.
  • Partner with ML researchers and engineers to understand how training loops, evaluation, and inference services consume data, and turn that into system requirements.
  • Define reliability, observability, and cost standards for data infrastructure, including SLOs, monitoring, capacity planning, and incident response.
  • Write design documents, review code, and mentor engineers on performance engineering and distributed systems practices.

Skills

Python
Rust
C++
Go

Education

BS/MS/PhD in Computer Science or related field

Tools

Spark
Ray
Dask
Kubernetes
Arrow
Parquet
Iceberg
Lance

Job description

Bayside Solutions, Inc. is seeking a Software Engineer to build and operate the data infrastructure feeding ML training and inference systems. You will design high-throughput, distributed data pipelines atop columnar formats to ensure data is readily available for GPUs and models.

The role emphasizes hands-on performance engineering, collaboration with ML researchers, and designing reliable, observable data infrastructure with cost awareness. W2 contract, remote-friendly in Cupertino, CA.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Software Engineer, ML Data Infrastructure
Software Engineer, ML Data Infrastructure

Bayside Solutions • Cupertino (CA)

Remote
USD 83,000 - 96,000
Remote ML Data Engineer for Large-Scale AI Pipelines
Remote ML Data Engineer for Large-Scale AI Pipelines

Bright-Vision-Technologies • United States

Remote
USD 100,000 - 150,000
Senior Remote ML Infrastructure Engineer
Senior Remote ML Infrastructure Engineer

BairesDev • Peru (IL)

On-site
USD 120,000 - 170,000
Remote work
Payment in USD
Home setup provided
+3
Remote ML Infra Engineer - High-Performance Inference
Remote ML Infra Engineer - High-Performance Inference

Bright Vision Technologies • Mountain View (CA)

On-site
USD 105,000 - 143,000
Remote ML Data Engineer - Scale-Power Data Pipelines
Remote ML Data Engineer - Scale-Power Data Pipelines

Bright Vision Technologies • Sterling (VA)

On-site
USD 100,000 - 150,000
Remote ML Data Engineer — Scale AI Data Pipelines
Remote ML Data Engineer — Scale AI Data Pipelines

Bright Vision Technologies • Folsom (CA)

Remote
USD 80,000 - 100,000
Remote ML Data Engineer - Petabyte-Scale Pipelines
Remote ML Data Engineer - Petabyte-Scale Pipelines

Bright Vision Technologies • Sterling (VA)

On-site
USD 100,000 - 150,000
Remote ML Platform Engineer: AI Infrastructure & Pipelines
Remote ML Platform Engineer: AI Infrastructure & Pipelines

SmartRecruiters, Inc. • San Francisco (CA)

Remote
USD 140,000 - 180,000
Medical Insurance
Dental Insurance
Vision Insurance
+2
Remote ML Data Engineer — Petabyte-Scale Pipelines
Remote ML Data Engineer — Petabyte-Scale Pipelines

Bright Vision Technologies • United States

Remote
USD 80,000 - 100,000
Staff Infra Engineer, ML Data Pipelines & GPUs
Staff Infra Engineer, ML Data Pipelines & GPUs

Sieve • San Francisco (CA)

On-site
USD 150,000 - 210,000
401k
Health insurance
Meals & snacks
+2