Senior System Engineer

Crystal Equation Corporation

New York (NY)

On-site

USD 140,000 - 200,000

Full time

5 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Crystal Equation Corporation is seeking an experienced infrastructure engineer to operate and scale Kubernetes platforms (EKS, GKE) across multiple clouds, manage cluster lifecycle, node pools, networking, and growth planning.

You will provision HPC infrastructure through CI/CD across AWS, CoreWeave, GCP, and OCI; work with Slurm for GPU scheduling; build monitoring, SLOs, and collaborate with Networking, Storage, Security, and AI/ML teams.

Qualifications

  • 4+ years in infrastructure engineering, cloud platforms, or HPC.
  • Terraform proficiency — writing and reviewing infrastructure-as-code daily.
  • Working knowledge of AWS (EC2, S3, EFS, FSx for Lustre).
  • Python for tooling and automation.
  • Slurm experience is a plus but not required.

Responsibilities

  • Operate and scale Kubernetes platforms (EKS, CKS, GKE) across multiple cloud providers — managing cluster lifecycle, node pools, networking, and growth planning.
  • Provision HPC infrastructure through CI/CD systems across AWS, CoreWeave, GCP, and OCI.
  • Work with Slurm-based job scheduling to allocate GPU compute for AI training and inference workloads.
  • Build and maintain monitoring, alerting, and SLOs — contributing to operational excellence and incident response.
  • Collaborate daily with Networking, Storage, Security, and AI/ML platform teams.

Skills

Infrastructure engineering
Cloud platforms
HPC
SRE practices

Tools

Terraform
AWS
Python
Slurm

Job description

  • Operate and scale Kubernetes platforms (EKS, CKS, GKE) across multiple cloud providers ?— managing cluster lifecycle, node pools, networking, and growth planning
  • Provision HPC infrastructure through CI/CD systems across AWS, CoreWeave, GCP, and OCI
  • Work with Slurm-based job scheduling to allocate GPU compute for AI training and inference workloads
  • Build and maintain monitoring, alerting, and SLOs ?— contributing to operational excellence and incident response
  • Collaborate daily with Networking, Storage, Security, and AI/ML platform teams
What We're Looking For
  • 4+ years in infrastructure engineering, cloud platforms, or HPC
  • Terraform proficiency — writing and reviewing infrastructure-as-code daily
  • Working knowledge of AWS (EC2, S3, EFS, FSx for Lustre)
  • Python for tooling and automation
  • Slurm experience is a plus but not required
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior System Engineer — Kubernetes, HPC & AI Infra
Senior System Engineer — Kubernetes, HPC & AI Infra

Crystal Equation Corporation • New York (NY)

On-site
USD 140,000 - 200,000
Staff Engineer, Senior Manager
Staff Engineer, Senior Manager

Jobtailor • Connecticut

On-site
USD 140,000 - 190,000
Platform Engineer - AI/ML Infrastructure (Kubernetes & Terraform)
Platform Engineer - AI/ML Infrastructure (Kubernetes & Terraform)

Madrona Venture Labs • United States

Hybrid
USD 180,000 - 260,000
Member of Technical Staff (AI Infrastructure Engineer)
Member of Technical Staff (AI Infrastructure Engineer)

Perplexity • San Francisco (CA)

On-site
USD 120,000 - 150,000
Principal AI Cloud Infra Architect: Kubernetes, Slurm, GPUs
Principal AI Cloud Infra Architect: Kubernetes, Slurm, GPUs

Jobtailor • California (MO)

On-site
USD 150,000 - 190,000
Senior Platform Engineer – Core Infrastructure
Senior Platform Engineer – Core Infrastructure

Jobtailor • California (MO)

On-site
USD 150,000 - 190,000
Senior Software Engineer
Senior Software Engineer

CoreWeave • Livingston (NJ)

On-site
USD 140,000 - 200,000
Kubernetes Platform Engineer
Kubernetes Platform Engineer

Ltd Global • Berkeley (CA)

On-site
USD 120,000 - 160,000
Platform Engineer
Platform Engineer

APN Consulting, Inc. • Jersey City (NJ)

On-site
USD 130,000 - 160,000
Senior Platform / DevOps Engineer
Senior Platform / DevOps Engineer

Veriipro • Westlake (TX)

On-site
USD 130,000 - 160,000