Senior Platform Engineer

STN Inc

Poland

Vor Ort

PLN 530.705 - 720.242

Vollzeit

14 Tage+
Bewerbungsgenerator

Hebe dich für diese Rolle von der Masse ab — erstelle in etwa einer Minute einen maßgeschneiderten Lebenslauf und ein Anschreiben.

Schaffe es an den ATS-Filtern vorbei

Zusammenfassung

STN Inc. is seeking a Senior Platform Engineer to design and operate the multi-tenant orchestration layer that turns raw GPU infrastructure into a usable cloud service. You will help build the backbone of GPUaaS with Kubernetes, Run:ai, and related tooling.

You will own multi-tenant isolation, platform APIs, image management, and automation, collaborating with SRE and TAM teams to ensure reliability and scalable growth. Remote US or hybrid in Pleasanton, CA options are supported.

Qualifikationen

  • 6+ years in platform engineering, SRE, or cloud engineering at scale.
  • Deep Kubernetes expertise including CRDs, operators, and multi-tenant patterns.
  • Strong programming skills in Go, Python, or both.
  • Experience operating GPU clusters or AI infrastructure at production scale.
  • Bachelor's degree in computer science or equivalent experience.

Aufgaben

  • Design and build the orchestration layer (Kubernetes, Slurm, Run:ai, or comparable)
  • Manage multi-tenant isolation including namespaces, networking, storage, and quotas
  • Build customer-facing platform APIs, CLIs, web portals, and SDKs
  • Implement and operate image management, GPU operator, and node provisioning automation
  • Drive infrastructure-as-code and automation across the platform stack
  • Partner with SRE on platform reliability, SLO definition, and observability
  • Support TAM and Support engineers on customer-impacting platform issues
  • Maintain customer environment templates, configuration management, and rollout tooling
  • Participate in architecture review, design discussions, and technical roadmap
  • Drive continuous platform improvement and reduce operational toil

Kenntnisse

Kubernetes
Go
Python

Ausbildung

Bachelor's degree in computer science

Tools

NVIDIA GPU Operator
MIG
MPS
NCCL operator
Slurm operator
Run:ai
KubeRay

Jobbeschreibung

Senior Platform Engineer

Platform and software · shared across customers

Reports to: Director, Platform Engineering (or Chief Architect)

Location: Remote (US) or Pleasanton, CA (hybrid)

Department: Cloud Platform Engineering / GPU Platform Engineering

Position summary

The Senior Platform Engineer builds and operates the multi-tenant orchestration, scheduling, and customer-facing platform layer that turns raw GPU infrastructure into a usable cloud service. This role is the software backbone of GPU One (GPUaaS).

Key responsibilities
  • Design and build the orchestration layer (Kubernetes, Slurm, Run:ai, or comparable)

  • Manage multi-tenant isolation including namespaces, networking, storage, and quotas

  • Build customer-facing platform APIs, CLIs, web portals, and SDKs

  • Implement and operate image management, GPU operator, and node provisioning automation

  • Drive infrastructure-as-code and automation across the platform stack

  • Partner with SRE on platform reliability, SLO definition, and observability

  • Support TAM and Support engineers on customer-impacting platform issues

  • Maintain customer environment templates, configuration management, and rollout tooling

  • Participate in architecture review, design discussions, and technical roadmap

  • Drive continuous platform improvement and reduce operational toil

Required qualifications
  • 6+ years in platform engineering, SRE, or cloud engineering at scale

  • Deep Kubernetes expertise including CRDs, operators, and multi-tenant patterns

  • Strong programming skills in Go, Python, or both

  • Experience operating GPU clusters or AI infrastructure at production scale

  • Bachelor's degree in computer science or equivalent experience

Preferred qualifications
  • Experience with NVIDIA GPU Operator, MIG, MPS, and NCCL operator patterns

  • Familiarity with Slurm operator, Run:ai, KubeRay, or comparable AI orchestration

  • Service mesh experience (Istio, Linkerd) and multi-cluster networking

  • Open source contributions in the cloud-native or AI infrastructure ecosystem

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.

oder ziehe deine Datei hierhin.

Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Senior GPU Platform Engineer — Remote
Senior GPU Platform Engineer — Remote

STN Inc • Polen

Hybrid
PLN 530.705 - 720.242
Technical Lead - GPU Infrastructure
Technical Lead - GPU Infrastructure

Tether • Warszawa

Vor Ort
PLN 300.000 - 520.000
Remote-friendly
Global collaboration
SW Platform Engineer (M/K)
SW Platform Engineer (M/K)

Experis ManpowerGroup Sp. z o.o. • Województwo pomorskie

Vor Ort
PLN 180.000 - 270.000
Senior AI Infrastructure & Platform Operations Engineer (remote in the EU)
Senior AI Infrastructure & Platform Operations Engineer (remote in the EU)

Mirantis • Poznań

Vor Ort
PLN 180.000 - 300.000
Remote Technical Operations & Deployment Engineer (GPU/AI/Cloud Infrastructure) at Radian Arc
Remote Technical Operations & Deployment Engineer (GPU/AI/Cloud Infrastructure) at Radian Arc

Radian Arc Limited • Polska

Remote
PLN 240.000 - 360.000
Senior Forward Deployed Solution Engineer (Poland) Engineering · Wroclaw · Onsite
Senior Forward Deployed Solution Engineer (Poland) Engineering · Wroclaw · Onsite

Spectro Cloud • Wrocław

Vor Ort
PLN 180.000 - 300.000
Hardware Engineer
Hardware Engineer

STN Inc • Polen

Vor Ort
PLN 454.890 - 682.335
Senior Cloud & HPC Architect — Kubernetes & GPU AI Infra
Senior Cloud & HPC Architect — Kubernetes & GPU AI Infra

NVIDIA • Warszawa

Vor Ort
PLN 221.000 - 507.000
Security and Compliance Engineer
Security and Compliance Engineer

STN Inc • Polen

Vor Ort
PLN 530.705 - 682.335
Senior Solutions Architect, DevOps
Senior Solutions Architect, DevOps

NVIDIA • Warszawa

Vor Ort
PLN 221.000 - 507.000