Contract Cloud SRE for Scalable GPU Platform

Sustainable Talent

Santa Clara (CA)

On-site

USD 89,544 - 117,096

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Full benefits
Paid time off
Company culture experience

Job summary

Sustainable Talent is partnering with Nvidia in Santa Clara for a Site Reliability Engineer (Contract) position. This full-time role offers pay ranging from $65/hr to $85/hr based on various factors and includes full benefits and paid time off.

The ideal candidate will have a strong background in maintaining and setting up Linux and Windows hosts, excellent debugging skills, and experience in scripting. This role involves collaborating within the cloud team to support infrastructure needs effectively.

Qualifications

  • 5+ years of experience in large-scale enterprise production systems.
  • Experience in debugging infrastructure issues.
  • Proficiency in scripting with Python or Go.

Responsibilities

  • Monitor and recover assets in the private cloud environment.
  • Deploy and maintain a large farm of machines.
  • Contribute to monitoring systems for real-time insights.

Skills

Debugging and analytical skills
Python or Go scripting
Linux and Windows systems maintenance

Education

Bachelor’s or Master’s Degree in Computer Science or Software Engineering

Tools

Chef
Ansible
Terraform
Version control systems such as Perforce or Git

Job description

Sustainable Talent is partnering with Nvidia in Santa Clara for a Site Reliability Engineer (Contract) position. This full-time role offers pay ranging from $65/hr to $85/hr based on various factors and includes full benefits and paid time off.

The ideal candidate will have a strong background in maintaining and setting up Linux and Windows hosts, excellent debugging skills, and experience in scripting. This role involves collaborating within the cloud team to support infrastructure needs effectively.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Sustainable Talent • Santa Clara (CA)

On-site
Full benefits
Paid time off
Company culture experience
Cloud SRE Architect — AI-Driven CI/CD & Scale
Cloud SRE Architect — AI-Driven CI/CD & Scale

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 272,000 - 431,000
Equity
Benefits
Senior GPU Cloud SRE — Kubernetes & Automation
Senior GPU Cloud SRE — Kubernetes & Automation

NVIDIA AI • Santa Clara (CA)

On-site
USD 168,000 - 334,000
Senior SRE: Global HPC & Multi-Cloud Reliability
Senior SRE: Global HPC & Multi-Cloud Reliability

NVIDIA Corporation • Durham (CA), Northern (KY)

Hybrid
USD 152,000 - 288,000
Senior Site Reliability Engineer - HPC
Senior Site Reliability Engineer - HPC

NVIDIA • United States

On-site
USD 184,000 - 288,000
Equity
Benefits package
Senior SRE: Automate Reliability for AI/GPU Platform
Senior SRE: Automate Reliability for AI/GPU Platform

Nscale • Seattle (WA)

On-site
USD 130,000 - 200,000
Lead Cloud SRE Architect for Private Cloud & AI CI/CD
Lead Cloud SRE Architect for Private Cloud & AI CI/CD

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 272,000 - 431,000
Equity
Benefits
Senior Site Reliability Engineer - HPC
Senior Site Reliability Engineer - HPC

NVIDIA Corporation • Durham (CA), Northern (KY)

On-site
USD 152,000 - 288,000
Senior Site Reliability Engineer - HPC
Senior Site Reliability Engineer - HPC

Socket.dev • Northern (KY)

Hybrid
USD 152,000 - 288,000
Equity
Benefits
Staff Site Reliability Engineer - AI Platform Runtime
Staff Site Reliability Engineer - AI Platform Runtime

Nvidia Corporation in • Santa Clara (CA)

Hybrid
USD 168,000 - 334,000