Infrastructure / DevOps Engineer

Discernis

New York (NY)

On-site

USD 150,000 - 190,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Discernis in New York is seeking an seasoned Infrastructure Engineer to own the platform infrastructure across cloud, customer managed cloud accounts and on-premises hardware, including air-gapped environments.

Responsibilities include packaging the platform as a reproducible Kubernetes-based bundle, operating GPU-enabled clusters, managing IaC and release processes, and building robust monitoring and security practices across diverse environments.

Qualifications

  • Experience with infrastructure as code tools such as Helm, Terraform, or Pulumi.
  • Hands on experience designing CI and CD workflows.
  • Comfort with observability stacks and production monitoring.
  • Strong networking and Linux fundamentals.
  • Bonus: GPU and AI workloads in Kubernetes, NVIDIA drivers and CUDA on bare metal, or shipping software into air gapped or customer controlled environments.
  • Bonus: Elasticsearch or similar distributed data systems.

Responsibilities

  • Package and ship our platform as a reproducible Kubernetes based bundle that installs cleanly on customer controlled and on premises hardware
  • Operate and improve Kubernetes clusters running both application services and GPU accelerated AI workloads
  • Own our infrastructure as code and release process so environments are reproducible across cloud and on prem
  • Build production grade monitoring, logging, and alerting that works even in restricted customer environments
  • Support large scale search and data infrastructure running on Kubernetes
  • Drive best practices in networking, secrets management, certificate handling, and security hardening

Skills

Helm
Terraform
Pulumi
CI/CD
Observability
Networking
Linux
GPU workloads
Elasticsearch

Tools

Kubernetes
Docker
NVIDIA drivers
Elasticsearch
K3s

Job description

Discernis builds AI driven document intelligence for high stakes legal work. Because our customers handle privileged and regulated matters that most cloud AI cannot touch, we run on premises and in customer controlled environments as well as in the cloud. That makes security, reliability, and reproducibility core product features, not afterthoughts. We work with AmLaw firms and enterprise legal teams where accuracy, explainability, and data control are non negotiable.

The Role

You will own the infrastructure that runs our platform in three very different places: our cloud, customer managed cloud accounts, and customer controlled hardware, including air gapped and on premises environments. The hard and interesting part of this job is making the same product deploy reliably and reproducibly on machines we do not operate. If you like turning bespoke deployments into a repeatable system, this is that job.

What You Will Do
  • Package and ship our platform as a reproducible Kubernetes based bundle that installs cleanly on customer controlled and on premises hardware
  • Operate and improve Kubernetes clusters running both application services and GPU accelerated AI workloads
  • Own our infrastructure as code and release process so environments are reproducible across cloud and on prem
  • Build production grade monitoring, logging, and alerting that works even in restricted customer environments
  • Support large scale search and data infrastructure running on Kubernetes
  • Drive best practices in networking, secrets management, certificate handling, and security hardening
What You Bring
  • Experience with infrastructure as code tools such as Helm, Terraform, or Pulumi
  • Hands on experience designing CI and CD workflows
  • Comfort with observability stacks and production monitoring
  • Strong networking and Linux fundamentals
  • Bonus: GPU and AI workloads in Kubernetes, NVIDIA drivers and CUDA on bare metal, or shipping software into air gapped or customer controlled environments
  • Bonus: Elasticsearch or similar distributed data systems
Tech Environment

Kubernetes and K3s, Docker, NVIDIA GPUs, Elasticsearch, Terraform and Helm, CI/CD tooling, cloud and on premises environments

Apply Here

https://app.dover.com/apply/Discernis%20AI/a4bc68ad-17e8-4a2c-877f-6b9d26ceb553?rs=42706078

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Kubernetes & AI Infra Engineer — On-Prem & Cloud
Kubernetes & AI Infra Engineer — On-Prem & Cloud

Discernis • New York (NY)

On-site
USD 150,000 - 190,000
Enterprise AI Platform Engineer
Enterprise AI Platform Engineer

DISCO • Austin (TX)

On-site
USD 120,000 - 150,000
Open, inclusive, and fun environment
Medical, dental, and vision insurance
401(k)
+3
Enterprise AI Platform Engineer
Enterprise AI Platform Engineer

DISCO • Atlanta (GA)

On-site
USD 120,000 - 150,000
Medical, dental and vision insurance
RSUs and competitive salary
Flexible PTO
Backend Engineer
Backend Engineer

Discernis • New York (NY)

On-site
USD 120,000 - 160,000
DevOps
DevOps

Complexio • Warsaw (IN)

On-site
USD 100,000 - 130,000
Opportunity for professional growth
Collaborative team environment
Continuous learning in a dynamic field
Infrastructure Engineer – DevOps, Kubernetes & Automation
Infrastructure Engineer – DevOps, Kubernetes & Automation

TensorWave • Las Vegas (NV)

On-site
USD 110,000 - 170,000
Stock Options
Medical Insurance
Dental Insurance
+9
Infrastructure Engineer – DevOps, Kubernetes & Automation
Infrastructure Engineer – DevOps, Kubernetes & Automation

TensorWave • Las Vegas (NM)

On-site
USD 80,000 - 110,000
Stock Options
100% paid Medical, Dental, and Vision insurance
Company Health Savings Account Contributions
+8
Software Engineer - Infrastructure
Software Engineer - Infrastructure

Emergentlabsinc • San Francisco (CA)

On-site
USD 110,000 - 150,000
401(k)
Health, dental, and vision insurance
Unlimited Paid Time Off
+1
Senior / Lead Infrastructure & Operations Engineer
Senior / Lead Infrastructure & Operations Engineer

Austin Werner • Boston (MA)

On-site
USD 120,000 - 150,000
AI Infrastructure & Platform Operations Engineer (remote in the US)
AI Infrastructure & Platform Operations Engineer (remote in the US)

Mirantis • United States

Remote
USD 110,000 - 150,000
Professional development
Conferences attendance
Team events