AWS Cloud Engineer for Scalable AI/HPC Infra

TensorWave

Las Vegas (NV)

On-site

USD 140,000 - 180,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Stock Options
Medical, Dental, and Vision insurance
HSA Contributions
Disability Insurance
Life Insurance
Pet & Legal Insurance
Supplementary Health Benefits
FSA
401(k)
Employee Assistance Program
Flexible PTO
Paid Holidays
Parental Leave
In-Office Perks

Job summary

TensorWave is hiring an AWS Cloud Engineer to design, provision, optimize, and support the AWS infrastructure powering our AMD GPU AI/HPC platform in the United States. This hands-on role collaborates with Rust backends, TypeScript developers, SREs, and platform teams to ensure reliable, cost-efficient cloud infrastructure that scales with demand.

You will own the full lifecycle of AWS infrastructure, implement IaC, and drive observability and security across environments.

Qualifications

  • 5+ years in cloud infrastructure, DevOps, SRE, or platform operations.
  • AWS experience with VPCs, EC2, S3, IAM, CloudWatch, Route 53, and load balancers.
  • Proficiency with IaC tooling; Terraform strongly preferred.
  • Strong Linux fundamentals; networking, storage, and troubleshooting.
  • Experience with CI/CD, Git workflows, and monitoring/alerting platforms.
  • Clear communicator able to document infrastructure across teams.

Responsibilities

  • Own full lifecycle of AWS infra across dev, staging, prod, and customer environments.
  • Build and maintain IaC with Terraform, Pulumi, AWS CDK, and CloudFormation.
  • Implement high availability, auto-scaling, and secure service communication.
  • Build and maintain CI/CD workflows for infra and hosted services.
  • Improve observability with metrics, logging, alerts, dashboards, and runbooks.
  • Troubleshoot AWS networking, compute, storage, IAM, and deployments.
  • Participate in incident response, post-incident reviews, and RCA.
  • Document architecture, operational processes, and best practices.

Skills

5+ yrs cloud infra
AWS experience
Terraform IaC
Linux fundamentals
CI/CD & monitoring
Documentation skills

Tools

Kubernetes (EKS)
Prometheus
Grafana
Datadog

Job description

TensorWave is hiring an AWS Cloud Engineer to design, provision, optimize, and support the AWS infrastructure powering our AMD GPU AI/HPC platform in the United States. This hands-on role collaborates with Rust backends, TypeScript developers, SREs, and platform teams to ensure reliable, cost-efficient cloud infrastructure that scales with demand.

You will own the full lifecycle of AWS infrastructure, implement IaC, and drive observability and security across environments.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

DevOps Engineer - AWS
DevOps Engineer - AWS

TensorWave • Las Vegas (NM)

On-site
USD 100,000 - 130,000
Stock Options
100% paid Medical, Dental, and Vision insurance
Flexible PTO
+1
DevOps Engineer - AWS
DevOps Engineer - AWS

TensorWave • Las Vegas (NV)

On-site
USD 140,000 - 180,000
Stock Options
Medical, Dental, and Vision insurance
HSA Contributions
+11
Senior AWS Cloud Engineer – AI/ML Infra & Observability
Senior AWS Cloud Engineer – AI/ML Infra & Observability

TensorWave • Las Vegas (NM)

On-site
USD 100,000 - 130,000
Stock Options
100% paid Medical, Dental, and Vision insurance
Flexible PTO
+1
Backend Platform Engineer – GPU Cluster Automation
Backend Platform Engineer – GPU Cluster Automation

TensorWave • Las Vegas (NV)

On-site
USD 140,000 - 210,000
Stock Options
Excellent health insurance
401(k)
+6
Lead Hardware Engineer - AI Cloud Infrastructure
Lead Hardware Engineer - AI Cloud Infrastructure

Amazon Web Services (AWS) • Austin (TX)

On-site
USD 159,000 - 215,000
Health insurance
401(k) matching
Paid time off
+1
AI Systems Engineer: HPC & GPU Clusters
AI Systems Engineer: HPC & GPU Clusters

Advanced Micro Devices, Inc. • San Jose (CA)

On-site
USD 180,000 - 260,000
AI Infrastructure Engineer - GPU & Kubernetes
AI Infrastructure Engineer - GPU & Kubernetes

HCLTech • California (MO)

On-site
USD 150,000 - 210,000
Medical Insurance
Dental Insurance
Vision Insurance
+2
Staff Software Engineer — AI Cloud Compute Platform
Staff Software Engineer — AI Cloud Compute Platform

Lambda Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 300,000
Health, dental, vision coverage
Wellness stipend
401k with 2% company match
+2
Staff Engineer, AI Cloud Infra (Kubernetes + GPUs)
Staff Engineer, AI Cloud Infra (Kubernetes + GPUs)

Lambda • San Francisco (CA)

Hybrid
USD 314,000 - 465,000
Health, dental, and vision coverage
401k with 2% company match
Wellness stipend
+1
AI Cloud Systems Engineer for Generative AI & ML Servers
AI Cloud Systems Engineer for Generative AI & ML Servers

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 99,000 - 160,000