Senior Platform Engineer — GPU Cloud Orchestration

STN Incorporated

United States

Hybrid

USD 140,000 - 180,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

STN Incorporated seeks a Senior Platform Engineer to design and operate the multi-tenant platform layer that turns GPU infrastructure into a usable cloud service. This role will own orchestration and automation across Kubernetes, Slurm, and Run:ai to support a scalable GPU cloud offering.

The position reports to the Director, Platform Engineering and requires 6+ years in platform engineering, deep Kubernetes expertise, and strong Go/Python coding skills. Remote US or hybrid in Pleasanton, CA.

Qualifications

  • 6+ years in platform engineering, SRE, or cloud engineering at scale.
  • Deep Kubernetes expertise including CRDs, operators, and multi-tenant patterns.
  • Strong programming skills in Go, Python, or both.
  • Experience operating GPU clusters or AI infrastructure at production scale.
  • Bachelor’s degree in computer science or equivalent experience.

Responsibilities

  • Design and build the orchestration layer (Kubernetes, Slurm, Run:ai, or comparable)
  • Manage multi-tenant isolation including namespaces, networking, storage, and quotas
  • Build customer-facing platform APIs, CLIs, web portals, and SDKs
  • Implement and operate image management, GPU operator, and node provisioning automation
  • Drive infrastructure-as-code and automation across the platform stack
  • Partner with SRE on platform reliability, SLO definition, and observability
  • Support TAM and Support engineers on customer-impacting platform issues
  • Maintain customer environment templates, configuration management, and rollout tooling
  • Participate in architecture review, design discussions, and technical roadmap
  • Drive continuous platform improvement and reduce operational toil

Skills

Go programming
Python programming
SRE / cloud engineering

Education

Bachelor's degree in computer science or equivalent

Tools

Kubernetes
NVIDIA GPU Operator
Slurm
Run:ai
KubeRay
Istio
Linkerd

Job description

STN Incorporated seeks a Senior Platform Engineer to design and operate the multi-tenant platform layer that turns GPU infrastructure into a usable cloud service. This role will own orchestration and automation across Kubernetes, Slurm, and Run:ai to support a scalable GPU cloud offering.

The position reports to the Director, Platform Engineering and requires 6+ years in platform engineering, deep Kubernetes expertise, and strong Go/Python coding skills. Remote US or hybrid in Pleasanton, CA.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Platform Engineer
Senior Platform Engineer

STN Incorporated • United States

Hybrid
USD 140,000 - 180,000
Senior GPU Platform Architect (GPUaaS)
Senior GPU Platform Architect (GPUaaS)

Unify Technologies Ltd • Plano (TX)

On-site
USD 180,000 - 240,000
Senior Platform Engineer - GPU-Driven Multi-Cloud Infra
Senior Platform Engineer - GPU-Driven Multi-Cloud Infra

Harrison Clarke • San Francisco (CA)

On-site
USD 120,000 - 160,000
Platform Engineer
Platform Engineer

Harrison Clarke • San Francisco (CA)

On-site
USD 120,000 - 160,000
Senior Kubernetes Platform Developer – GPU & AI Infrastructure
Senior Kubernetes Platform Developer – GPU & AI Infrastructure

GTN Technical Staffing • Dallas (TX)

Hybrid
USD 165,000 - 210,000
Relocation available
Hybrid work model
Senior Cloud-Native Engineer — Kubernetes & Slurm for Multi-Tenant GPUs
Senior Cloud-Native Engineer — Kubernetes & Slurm for Multi-Tenant GPUs

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 184,000 - 287,500
Equity options
Comprehensive benefits
Platform Engineer - GPU Cloud Control Plane
Platform Engineer - GPU Cloud Control Plane

Hyperbolic • San Francisco (CA)

On-site
USD 140,000 - 210,000
Senior Cloud Platform Engineer: Core GPU Infrastructure
Senior Cloud Platform Engineer: Core GPU Infrastructure

Lambda • San Francisco (CA)

Hybrid
USD 296,000 - 346,000
Health insurance
Dental insurance
Vision insurance
+2
Kubernetes Platform Engineer - GPU & AI Infra
Kubernetes Platform Engineer - GPU & AI Infra

GTN Technical Staffing • Dallas (TX)

Hybrid
USD 165,000 - 210,000
Relocation available
Hybrid work model
Senior Backend Engineer, GPU Cloud Infra & Kubernetes
Senior Backend Engineer, GPU Cloud Infra & Kubernetes

Socket.dev • New York (NY)

Hybrid
USD 180,000 - 250,000
Health insurance
Equity
401(k) matching
+6