Enable job alerts via email!

Senior DevOps Infrastructure Engineer, Open-Source CI and CD

NVIDIA

New Jersey

Remote

USD 168,000 - 334,000

Full time

8 days ago

Boost your interview chances

Create a job specific, tailored resume for higher success rate.

Job summary

An innovative company is seeking a Senior DevOps Infrastructure Engineer to enhance their open-source CI/CD processes. This role involves managing self-hosted GitHub Actions runners and optimizing infrastructure using Kubernetes and Terraform. You will collaborate with a talented team to support developers and ensure a seamless CI/CD experience. If you thrive in a remote environment and have a passion for cutting-edge technology, this is your chance to contribute to impactful projects in the AI and computing space. Join a forward-thinking team that values creativity and autonomy in a diverse workplace.

Qualifications

  • 7+ years of experience in infrastructure, DevOps, or platform engineering.
  • Strong expertise in Kubernetes and GitOps tools like ArgoCD.
  • Proficiency in Linux administration and troubleshooting.

Responsibilities

  • Manage and scale self-hosted GitHub Actions runners using Kubernetes.
  • Use Infrastructure as Code to deploy and maintain infrastructure.
  • Develop and deploy custom Golang tools for platform observability.

Skills

Kubernetes
Infrastructure as Code (Terraform)
Golang
Python
Linux Administration
GitOps (ArgoCD)
Monitoring (Prometheus, Grafana)
CI/CD Pipelines

Education

B.S. or M.S. in Computer Science

Tools

HashiCorp Packer
GitHub Actions

Job description

Senior DevOps Infrastructure Engineer, Open-Source CI and CD

Join to apply for the Senior DevOps Infrastructure Engineer, Open-Source CI and CD role at NVIDIA

Senior DevOps Infrastructure Engineer, Open-Source CI and CD

Join to apply for the Senior DevOps Infrastructure Engineer, Open-Source CI and CD role at NVIDIA

At NVIDIA, we are pushing the boundaries of AI, graphics, and computing. The GitHub Actions Runner team manages self-hosted GPU-enabled GitHub Actions runners, using Actions Runner Controller with KubeVirt to deploy ephemeral VM-based runners for NVIDIA’s open source projects on GitHub. The team operates both on-premise and in the cloud (AWS) to support 100+ developers with whom they collaborate regularly to ensure a seamless CI/CD experience. We are looking for a Senior Infrastructure Engineer to help scale, optimize, and expand our platform.

Our team is fully remote and distributed across multiple time zones. If you're passionate about infrastructure, Kubernetes, automation, and observability, this is an opportunity to work with exciting technology at one of the most innovative companies in the world.

Preferred work location: Eastern/Central time zones

What You'll Be Doing

  • Manage and scale self-hosted GitHub Actions runners using Kubernetes
  • Help expand runner support for various hardware and operating system combinations, including Linux, Windows, single-GPU, multi-GPU, NVLink, and more
  • Use Infrastructure as Code (Terraform and ArgoCD) to deploy and maintain infrastructure both on-premise and in AWS
  • Build and maintain runner VM images using HashiCorp Packer
  • Connect distributed services securely using mTLS, PKI, and HashiCorp Vault
  • Develop, package, and deploy custom Golang tools to support platform observability, stability, and efficiency
  • Configure alerting and monitoring to identify and address issues quickly, using tools like Prometheus and Grafana
  • Contribute upstream to open-source tools and libraries that our team depends on
  • Periodically update platform dependencies and address CVEs

What We Need To See

  • B.S. or M.S. in Computer Science, Computer Engineering, or a related field (or equivalent experience)
  • 7+ years of proven experience in infrastructure, DevOps, or platform engineering
  • Strong Kubernetes expertise (running, debugging, and scaling workloads)
  • Experience with GitOps tools (ArgoCD or similar)
  • Proficiency in Linux administration and troubleshooting
  • Experience with Infrastructure as Code using Terraform/Terragrunt
  • Proficiency in Golang, Python, and TypeScript
  • Hands-on experience with monitoring, logging, and tracing (Prometheus, Grafana, OpenTelemetry, etc.)
  • Solid understanding of CI/CD pipelines, particularly GitHub Actions
  • Ability to work and collaborate effectively with a fully remote, distributed team

Ways To Stand Out From The Crowd

  • Experience instrumenting telemetry for distributed systems
  • Strong background in GPU workloads on Kubernetes with experience writing custom Kubernetes controllers
  • Deep understanding of KubeVirt and/or virtualization
  • Experience with self-hosted GitHub Actions runners
  • Contributions to open-source Kubernetes-related projects

With competitive salaries and a generous benefits package, NVIDIA is considered one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking individuals in the industry working for us. Due to unprecedented growth, our exclusive engineering teams are expanding rapidly. If you're a creative and autonomous engineer with a genuine passion for technology, we want to hear from you!

The base salary range is 168,000 USD - 333,500 USD. Your base salary will be determined based on your location, experience, and the pay of employees in similar positions.

You will also be eligible for equity and benefits . NVIDIA accepts applications on an ongoing basis.

NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

JR1996712

Seniority level
  • Seniority level
    Mid-Senior level
Employment type
  • Employment type
    Full-time
Job function
  • Job function
    Information Technology
  • Industries
    Computer Hardware Manufacturing, Software Development, and Computers and Electronics Manufacturing

Referrals increase your chances of interviewing at NVIDIA by 2x

Sign in to set job alerts for “Senior Infrastructure Engineer” roles.
System Administrator with Linux/ Python / R - Language Experience - Remote

Parsippany, NJ $90,000.00-$120,000.00 1 week ago

Senior Software Engineer – AI Infrastructure and Tooling

Newark, NJ $200,000.00-$300,000.00 1 week ago

We’re unlocking community knowledge in a new way. Experts add insights directly into each article, started with the help of AI.

Get your free, confidential resume review.
or drag and drop a PDF, DOC, DOCX, ODT, or PAGES file up to 5MB.

Similar jobs

Senior DevOps Infrastructure Engineer, Open-Source CI and CD

NVIDIA

Remote

USD 168,000 - 334,000

8 days ago

Senior/Staff Engineer, Infrastructure (DevOps)

Pryon

Washington

Remote

USD 180,000 - 215,000

Yesterday
Be an early applicant

Senior Software Engineer, Distributed Systems

Censys, Inc.

Kirkland

Remote

USD 149,000 - 190,000

Yesterday
Be an early applicant

Senior Software Engineer, Distributed Systems

Censys, Inc.

Los Altos

Remote

USD 149,000 - 190,000

Yesterday
Be an early applicant

Senior Software Engineer - Distributed Systems & File Sync

Air

Remote

USD 160,000 - 264,000

2 days ago
Be an early applicant

Senior Software Engineer, Distributed Systems

Censys

Tysons

Remote

USD 149,000 - 190,000

2 days ago
Be an early applicant

Senior Software Engineer (Infrastructure Engineer)

ZipRecruiter

Seattle

Remote

USD 160,000 - 180,000

8 days ago

Senior Software Engineer (Infrastructure Engineer)

Textio

Seattle

Remote

USD 160,000 - 180,000

10 days ago

Staff Software Engineer (Infrastructure Platform)

Affirm

Remote

GBP 140,000 - 180,000

25 days ago