Senior Compute Support Engineer: Linux & Kubernetes

Mistral

San Francisco (CA)

Hybrid

USD 140,000 - 190,000

Full time

2 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Healthcare coverage
Relocation support
Wellness programs

Job summary

Mistral is building a Compute Support team to ensure reliability, performance, and scalability of GPU clusters in a Kubernetes environment. You will be the first contact for customers and internal teams, providing guidance and troubleshooting for compute-related inquiries in a hybrid role.

As a founding member, you will shape processes, standards, and culture, with on-call rotations and collaboration with SRE teams to optimize IaC and monitoring.

Qualifications

  • 5+ years of experience in system administration, technical support, or infrastructure operations in compute-heavy environments.
  • Deep Linux/Unix expertise with kernel debugging and performance tuning.
  • Hands-on Kubernetes experience with pod/node troubleshooting and networking/storage concepts.
  • Familiarity with container runtimes (Docker, containerd) and basic IaC tooling.
  • Experience with monitoring (Prometheus, Grafana) and log analysis.

Responsibilities

  • Serve as L1/L2 support for Linux, Kubernetes, and compute infrastructure issues.
  • Diagnose performance bottlenecks, hardware failures, and resource contention in distributed systems.
  • Triage requests, provide actionable guidance, and maintain runbooks and post-mortems.
  • Collaborate with SRE teams to improve IaC and tooling (Go-based).
  • Participate in on-call rotations to ensure 24/7 coverage.

Skills

Linux/Unix
Kernel debugging
Kubernetes
Containerization
Docker
Go
Python/Bash
Monitoring tools
On-call / SLAs
System performance tuning

Tools

Prometheus
Grafana
ELK/OpenTelemetry
Terraform/Ansible

Job description

Mistral is building a Compute Support team to ensure reliability, performance, and scalability of GPU clusters in a Kubernetes environment. You will be the first contact for customers and internal teams, providing guidance and troubleshooting for compute-related inquiries in a hybrid role.

As a founding member, you will shape processes, standards, and culture, with on-call rotations and collaboration with SRE teams to optimize IaC and monitoring.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Compute Support Engineer - Kubernetes & Linux Expert
Compute Support Engineer - Kubernetes & Linux Expert

Lindus Health • Paris (TX)

On-site
USD 120,000 - 180,000
Healthcare coverage
Relocation support
Meal and transportation allowances
Technical Support Engineer (L2) - Compute
Technical Support Engineer (L2) - Compute

Mistral • San Francisco (CA)

Hybrid
USD 140,000 - 190,000
Healthcare coverage
Relocation support
Wellness programs
AI Infrastructure Engineer — Kubernetes, Linux & Cloud
AI Infrastructure Engineer — Kubernetes, Linux & Cloud

Lindus Health • Palo Alto (CA)

On-site
USD 140,000 - 200,000
ML Platform Engineer — Scalable GPU & Kubernetes
ML Platform Engineer — Scalable GPU & Kubernetes

Socket.dev • Palo Alto (CA)

On-site
USD 180,000 - 280,000
Healthcare coverage
Relocation support
Retirement plans
+2
ML Platform Engineer: Scalable GPU + Kubernetes
ML Platform Engineer: Scalable GPU + Kubernetes

Mistral • Palo Alto (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Healthcare coverage
Parental leave
Relocation support
+2
Sovereign AI Compute Engineer | Cloud & Kubernetes
Sovereign AI Compute Engineer | Cloud & Kubernetes

Mistral • Palo Alto (CA)

On-site
USD 150,000 - 210,000
Healthcare coverage
Parental leave
Retirement plans
+3
AI Compute Solutions Architect
AI Compute Solutions Architect

Engg • New York (NY)

On-site
USD 150,000 - 230,000
Healthcare coverage
Parental leave
Retirement plans
+4
Senior Kubernetes & GPU Infra Engineer for AI-scale Compute
Senior Kubernetes & GPU Infra Engineer for AI-scale Compute

Kindredventures • United States

On-site
USD 140,000 - 190,000
ML Platform Engineer: GPU Orchestration & Scale
ML Platform Engineer: GPU Orchestration & Scale

Mistral • Palo Alto (CA)

On-site
USD 180,000 - 280,000
Healthcare coverage
Parental leave
Retirement plans
+3
Senior Cloud Support Engineer — Kubernetes & GPU
Senior Cloud Support Engineer — Kubernetes & GPU

Vultr • Town of Montana (WI)

On-site
USD 90,000 - 110,000
Company-paid health premiums
401(k) with matching
Professional development reimbursement
+6