Site Reliability Engineer — Remote, NixOS & Observability

Leidos Inc

San Antonio (TX)

On-site

USD 87,000 - 157,000

Full time

12 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Flexible work hours
Work-from-home options
Home-roasted coffee and award-winning?
Office amenities and collaboration

Job summary

Leidos is seeking a Site Reliability Engineer to help build reliable compute platforms and Nix/NixOS-based infrastructure. You will design, deploy, and automate Linux environments, create reusable modules and deployment workflows, and drive resilience and observability across distributed systems.

The role emphasizes declarative systems, infrastructure-as-code, and reducing configuration drift. You will work on high-performance compute, distributed storage, networking, and ML/AI infrastructure,

Qualifications

  • Bachelor's degree or equivalent experience in computer science or related field.
  • Strong Linux system administration and troubleshooting experience.
  • Experience with Linux networking, storage, and file systems.
  • Proficient in automating system configuration and deployment.
  • Experience with Nix/NixOS in production or labs.
  • 2+ years of scripting in Python, Go, or Bash.
  • Ability to debug complex systems methodically across layers.

Responsibilities

  • Own the reliability and operation of critical compute and data platforms.
  • Design, deploy, and maintain Linux infrastructure, including NixOS-based systems.
  • Build reproducible configurations and deployment workflows using Nix and IaC.
  • Develop reusable NixOS modules, packages, flakes, and system configurations.
  • Improve platform resilience, fault tolerance, recoverability, and maintainability.
  • Build monitoring, logging, alerting, and observability to preempt issues.
  • Diagnose failures across Linux, networking, storage, containers, and distributed systems.
  • Automate repetitive operational tasks and eliminate drift.
  • Contribute to security and maintainability of deployed systems.
  • Help define deployment practices, upgrades, and disaster-recovery procedures.
  • Perform capacity planning and identify performance bottlenecks.

Skills

Enthusiasm for learning
Linux systems administration
Linux networking
Automation of configuration
Scripting (Python/Go/Bash)
Debugging complex systems

Education

Bachelor's degree in Computer Science/Engineering or related field

Tools

Nix/NixOS
Terraform
Ansible

Job description

Leidos is seeking a Site Reliability Engineer to help build reliable compute platforms and Nix/NixOS-based infrastructure. You will design, deploy, and automate Linux environments, create reusable modules and deployment workflows, and drive resilience and observability across distributed systems.

The role emphasizes declarative systems, infrastructure-as-code, and reducing configuration drift. You will work on high-performance compute, distributed storage, networking, and ML/AI infrastructure,

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer: NixOS, Observability, Remote
Site Reliability Engineer: NixOS, Observability, Remote

Via Logic LLC • San Antonio (TX)

On-site
USD 87,000 - 157,000
Site Reliability Engineer — NixOS, Monitoring & Resilience
Site Reliability Engineer — NixOS, Monitoring & Resilience

Leidos Inc • Columbus (OH)

Hybrid
USD 87,000 - 157,000
Work-from-home options
Site Reliability Engineer (NixOS & Infra Automation)
Site Reliability Engineer (NixOS & Infra Automation)

Leidos • San Antonio (TX)

Hybrid
USD 87,000 - 157,000
Health and Wellness programs
Retirement plan
Paid Leave
Site Reliability Engineer — NixOS, Remote & Flexible Hours
Site Reliability Engineer — NixOS, Remote & Flexible Hours

Leidos Inc • Boulder (CO)

Hybrid
USD 87,000 - 157,000
NixOS Site Reliability Engineer
NixOS Site Reliability Engineer

Leidos • Boulder (CO)

Hybrid
USD 87,000 - 157,000
Health and Wellness programs
Paid Leave
Retirement
Site Reliability Engineer - Remote, NixOS & Resilient Infra
Site Reliability Engineer - Remote, NixOS & Resilient Infra

Leidos • Columbus (OH)

Hybrid
USD 87,000 - 157,000
Flexible work hours
Work-from-home options
Coffee in office
Site Reliability Engineer — NixOS & Infra-as-Code Lead
Site Reliability Engineer — NixOS & Infra-as-Code Lead

Via Logic LLC • Boulder (CO)

Hybrid
USD 87,000 - 157,000
NixOS SRE: Reliable Infra, Reproducible Deployments Remote
NixOS SRE: Reliable Infra, Reproducible Deployments Remote

Leidos • Chantilly (VA)

Hybrid
USD 87,000 - 157,000
Flexible work hours
Work-from-home options
Office coffee
+1
NixOS SRE: Build Reproducible, Resilient Infra
NixOS SRE: Build Reproducible, Resilient Infra

Via Logic LLC • Chantilly (VA)

On-site
USD 87,000 - 157,000
Remote Site Reliability Engineer: NixOS & Infra Automation
Remote Site Reliability Engineer: NixOS & Infra Automation

Leidos Inc • Chantilly (VA)

On-site
USD 87,000 - 157,000