Senior Backend Engineer - Scalable ML Inference (Golang)

Unity

Mountain View (CA)

On-site

USD 210,000 - 273,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Health insurance
Equity awards
Retirement/pension plans
Generous vacation and personal days
Parental leave support

Job summary

Unity is hiring for a backend engineering role focused on production-grade services powering online model inference at scale. You will design, deploy, and optimize low-latency, high-throughput infrastructure on GCP with Kubernetes, while partnering with ML engineers to handle increasing model complexity.

Strong observability with Prometheus/Grafana is essential. The role emphasizes DevOps/SRE practices, Docker-based microservices, and exposure to ML serving platforms.

Qualifications

  • Experience designing, deploying, and maintaining distributed systems at scale.
  • Expertise in Golang for high-performance backend infrastructure.
  • Hands-on experience with GCP and workload orchestration using Kubernetes.
  • Solid grounding in monitoring/observability with Prometheus and Grafana.
  • Experience in ad tech, recommender systems, real-time personalization, or similar domains.
  • Familiarity with microservice architectures, Docker, and CI/CD practices.
  • Familiarity with ML platforms, workflows, and serving infrastructure.

Responsibilities

  • Design, develop, and deploy production-grade backend services and distributed systems.
  • Lead technical direction of inference platform for low-latency serving.
  • Collaborate with ML engineers to scale online serving with growing models.
  • Ensure reliability and efficiency using monitoring/observability tools.
  • Manage cloud infrastructure on GCP and orchestrate workloads with Kubernetes.
  • Promote best practices for backend development, testing, deployment, and monitoring.

Skills

Golang
Distributed systems
GCP
Kubernetes
Prometheus
Grafana
Docker
CI/CD
ML platforms

Tools

Terraform
Mesos/K8s tooling

Job description

Unity is hiring for a backend engineering role focused on production-grade services powering online model inference at scale. You will design, deploy, and optimize low-latency, high-throughput infrastructure on GCP with Kubernetes, while partnering with ML engineers to handle increasing model complexity.

Strong observability with Prometheus/Grafana is essential. The role emphasizes DevOps/SRE practices, Docker-based microservices, and exposure to ML serving platforms.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior ML Infrastructure Engineer - Online Inference
Senior ML Infrastructure Engineer - Online Inference

Unity Enterprise • Northern (KY)

Hybrid
USD 210,000 - 273,000
Comprehensive health insurance
Employee stock ownership
Generous vacation and personal days
+1
Senior ML Infra Engineer — Production Inference Architect
Senior ML Infra Engineer — Production Inference Architect

LE130 Unity Technologies SF • United States

Remote
USD 210,000 - 273,000
Health insurance
Stock options
Retirement plans
+2
Senior ML Infra Engineer: Production Inference
Senior ML Infra Engineer: Production Inference

Unity • United States

Remote
USD 187,000 - 260,000
Health insurance
Stock options
Retirement plans
+1
Senior ML Infra Engineer — Real-Time Inference
Senior ML Infra Engineer — Real-Time Inference

Unity Technologies SF • Olympia (WA)

On-site
USD 166,000 - 273,000
Comprehensive health insurance
Commute subsidy
Employee stock ownership
+3
Staff Backend Engineer, ML Inference Systems
Staff Backend Engineer, ML Inference Systems

Unity • Mountain View (CA)

On-site
USD 245,000 - 318,000
Health insurance
Stock ownership
Commute subsidy
+4
Cloud-Scale Backend Engineer for ML Inference
Cloud-Scale Backend Engineer for ML Inference

Praxis, Inc. • San Francisco (CA)

On-site
USD 170,000 - 250,000
ML Inference Platform Engineer — Scale Production AI
ML Inference Platform Engineer — Scale Production AI

The Consensus • New York (NY)

On-site
USD 120,000 - 150,000
100% coverage of medical, dental, and vision insurance
Flexible PTO policy
Paid parental leave
+2
Remote Senior ML Infra Engineer — Real-Time Model Serving
Remote Senior ML Infra Engineer — Real-Time Model Serving

3M HEALTHCARE • Mountain View (CA)

On-site
USD 183,700 - 248,600
Health insurance
Stock options
Retirement plan
+1
Remote Senior ML Infra Engineer — Real-Time Model Serving
Remote Senior ML Infra Engineer — Real-Time Model Serving

3M HEALTHCARE • Mountain View (CA)

On-site
USD 183,700 - 248,600
Health insurance
Stock options
Retirement plan
+1
Senior Real-Time ML Infrastructure Engineer
Senior Real-Time ML Infrastructure Engineer

3M HEALTHCARE • Bellevue (WA)

Remote
USD 183,700 - 248,600