DATACENTER Engineer

Sustainable Talent

California

On-site

USD 110,208 - 123,984

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

401(k)
Vision insurance
Medical insurance
Full benefits
PTO
Amazing company culture

Job summary

Join an innovative firm as a Data Center Engineer, where you will support a cutting-edge private cloud infrastructure team. This exciting role involves maintaining a compute farm that serves as a test-bed for advanced technology, ensuring high availability and reliability. You will collaborate with talented engineers to drive efficiency and implement improvements while managing critical data center operations. With a focus on teamwork and problem-solving, this position offers a dynamic environment where your contributions will directly impact the success of the organization. Embrace the opportunity to work in a culture that values creativity and excellence!

Qualifications

  • 5+ years of experience in data centers or large engineering labs.
  • Proficiency in scripting (shell, Python, Ansible) and DCIM tools.

Responsibilities

  • Manage and maintain a high-performing Compute Farm of builders, packagers, and testers.
  • Collaborate with engineering teams to design and release next-generation products.

Skills

Problem-solving skills
Communication skills
Collaboration
Technical curiosity

Education

Associate’s or Bachelor’s Degree in Engineering/Technical Major

Tools

GIT
Perforce
DCIM (Nautobot)
Python
Ansible
Windows
Linux
Mac OS

Job description

Sustainable Talent is partnering with Nvidia, a global leader that's been transforming computer graphics, PC gaming, and accelerated computing for over 25 years.

We are looking for a Data Center Engineer to support our client's on-premise, private cloud infrastructure team. This is a W-2 full-time contract based in Santa Clara, CA. We offer competitive pay $80-$90/hr based on factors like experience, education, location, etc. and provide full benefits, PTO, and amazing company culture!

In this role, you will be faced with the challenge of providing and maintaining a compute farm of systems which includes Builders, Packagers, and Testers that act as a test-bed for our developers worldwide to test various Nvidia hardware and software prior to release. The environment is huge, the scale massive, and the ask enormous! We need YOU to help US maintain and drive our world-class DCs/Labs to produce timely, deterministic results for our Engineers and expectant Users worldwide!

What You'll Do:
  • Collaborate closely with engineering teams (system architects, hardware/software engineers, QA, and more) to design, develop, debug, and release next-generation products.
  • Manage and maintain a high-performing Compute Farm of builders, packagers, testers, and core infrastructure.
  • Ensure availability targets are consistently met and lead system recovery efforts.
  • Deploy and qualify systems while supporting exciting new technology bring-ups.
  • Oversee inventory and lifecycle management for NVIDIA's assets across data centers and labs.
  • Gather critical metrics and create Standard Operating Procedures (SOPs) documentation.
  • Maintain a world-class, safe, and well-organized environment in our data centers and labs.
  • Troubleshoot Linux/Windows, hardware, and infrastructure issues alongside engineers and platform operations teams.
  • Plan, deploy, and maintain on-premises private cloud infrastructure, collaborating with datacenter and network engineering teams.
  • Implement efficiency improvements to maximize availability, throughput, and test accuracy while meeting SLAs and KPIs.
  • Represent the team in meetings with internal stakeholders and contribute to global operations.
What We Need to See:
  • Associate’s or Bachelor’s Degree in Engineering/Technical Major (or equivalent experience).
  • 5+ years of experience in data centers or large engineering labs.
  • Familiarity with SCMs like GIT/Perforce.
  • Proficiency in DCIM (Nautobot, etc.) and scripting (shell, Python, Ansible).
  • Working knowledge of protocols/services like TCP/IP, DNS, NFS, SSL, etc.
  • Experience with Windows, Linux, and Mac operating systems.
  • Hands-on experience with PCBs, GPUs, and system deployments.
  • Exceptional communication skills, both written and verbal.
  • Ability to explain technical concepts to non-technical audiences.
  • Strong problem-solving skills and a collaborative spirit.
What Makes You Stand Out:
  • Experience managing HPC clusters using tools like BCM and Slurm.
  • Hands-on knowledge of OpenStack.
  • Relevant certifications such as CCNA or equivalent.
  • Strong background in Windows and Linux administration, with an understanding of dense datacenter design, including compute, storage, and networking.
  • Experience with hypervisors and VM applications.
  • Knowledge of DC infrastructure with an emphasis on liquid cooling.
  • A track record of technical curiosity and innovation.
  • Mechanically inclined and comfortable with tools and physical tasks.
  • Energetic, enthusiastic, and the understanding of what it takes to get the team to the finish line.
  • Willing to go the extra mile to get the job done!
  • This is an onsite contract position, and will require local travel to DCs within Santa Clara.

Sustainable Talent is a M/F+, disabled, and veteran equal employment opportunity and affirmative action employer.

Seniority level

Mid-Senior level

Employment type

Contract

Job function

Referrals increase your chances of interviewing at Sustainable Talent by 2x

Inferred from the description for this job

401(k)

Vision insurance

Medical insurance

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Systems Administrator supporting Nvidia
Systems Administrator supporting Nvidia

Sustainable Talent • Santa Clara (CA)

On-site
USD 143,270,000 - 229,233,000
Full benefits
PTO
Culture of excellence
DevOps Engineer
DevOps Engineer

Insight Global • California

On-site
Medical Insurance
Dental Insurance
Vision Insurance
+4
Senior Manager, Business Operations
Senior Manager, Business Operations

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 240,000 - 380,000
Solutions Architect, Data Center Buildouts
Solutions Architect, Data Center Buildouts

NVIDIA Gruppe • Town of Texas (WI)

On-site
USD 224,000 - 431,000
Equity
Benefits
Senior Data Center Infrastructure Engineer
Senior Data Center Infrastructure Engineer

2100 NVIDIA USA • California (MO)

On-site
USD 168,000 - 265,000
Equity
Benefits
Solutions Architect, Data Center Buildouts
Solutions Architect, Data Center Buildouts

NVIDIA • California (MO)

On-site
USD 224,000 - 356,500
Equity
Benefits
Systems Operations and Administrator
Systems Operations and Administrator

NVIDIA AI • Santa Clara (CA)

On-site
USD 112,000 - 219,000
Equity
Data Center Compute Engineer
Data Center Compute Engineer

Blue Signal Search • San Francisco (CA)

Hybrid
USD 150,000 - 210,000
Competitive compensation
Equity opportunity
Comprehensive benefits
+2
Data Center Compute Engineer
Data Center Compute Engineer

Blue Signal Search • United States

Hybrid
USD 120,000 - 180,000
Competitive compensation
Equity opportunity
Comprehensive benefits
Senior Technical Program Manager, Data Center - Engineering Operations
Senior Technical Program Manager, Data Center - Engineering Operations

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 168,000 - 258,750
Equity
Comprehensive benefits