Site Reliability Engineer

Leidos

Deutschland

Hybrid

EUR 80.000 - 145.000

Vollzeit

Vor 8 Tagen

Erhalte mehr Antworten von Arbeitgebern

Versende in nur wenigen Minuten einen passgenauen Lebenslauf.

Zusammenfassung

Kudu Dynamics is seeking a Site Reliability Engineer to build and operate reliable, reproducible compute and data platforms. You will focus on Nix/NixOS, automation, and IaC across Linux systems, networking, storage, and deployment pipelines.

Join a team that values declarative configurations, infrastructure-as-code, and reducing configuration drift, while delivering high-performance and secure infrastructure in a flexible work environment.

Qualifikationen

  • Bachelor's degree or equivalent in CS/CE or related field.
  • 2+ years Python/Go/Bash scripting experience.
  • Strong Linux administration and troubleshooting skills.
  • Experience with Nix/NixOS in production or significant labs.
  • Experience with Linux networking, storage, and file systems.
  • Experience with automation of configuration and deployment.
  • Ability to debug complex systems across stack layers.

Aufgaben

  • Own reliability and operation of critical compute and data platforms.
  • Design, deploy, and maintain Linux infrastructure including NixOS-based systems.
  • Build reproducible configurations and deployment workflows using Nix and IaC.
  • Develop reusable NixOS modules, packages, flakes, and configurations.
  • Improve platform resilience, fault tolerance, and maintainability.
  • Build monitoring, logging, alerting, observability to understand system behavior.
  • Diagnose failures across Linux, networking, storage, containers, and distributed systems.
  • Automate repetitive tasks and reduce configuration drift.
  • Design security and maintainability components for deployed systems.
  • Define operational standards, deployment practices, and disaster recovery.
  • Perform capacity planning and identify performance bottlenecks.

Kenntnisse

Linux administration
Networking
Storage systems
Automation
Nix/NixOS
Scripting (Python/Go/Bash)
Debugging complex systems

Ausbildung

Bachelor's degree in Computer Science or related field

Tools

NixOS
Terraform
Ansible
Prometheus
Grafana
Loki
OpenTelemetry
Elasticsearch

Jobbeschreibung

Description
Who We Are:

Kudu Dynamics is a 100% employee-owned company, forged out of a decade of experience in computer network operations and staffed with talent who have built, overseen, and enhanced capabilities throughout the entire USG arsenal. Our team of hackers, engineers, makers, and shakers brings deep experience across research, development, deployment, and operations.

Kudu Dynamics is uniquely qualified to anticipate tomorrow's threats and build the next generation of capabilities.

When you come in for your interview, you'll see that Kudu is an amazing place to work, where you'll be surrounded by experts who are ready to teach and learn. Our team has flexible work hours and work-from-home options. When we do work from the offices, we enjoy home-roasted coffee and award-winning workspaces.

Job Description:

Hello! We're a small team in a fun company looking for someone to come in and help us build systems that stay reliable when things get complicated.

We need a Site Reliability Engineer who has experience building, deploying, automating, and operating complex compute platforms. You'll work across infrastructure, Linux systems, networking, distributed storage, observability, security, and deployment automation, with a particular emphasis on creating systems that are reproducible, maintainable, resilient, and easy to operate.

A major part of this role will involve Nix and NixOS. We want someone who appreciates declarative systems, reproducible environments, infrastructure-as-code, and reducing configuration drift. You may be building NixOS-based servers, improving deployment pipelines, developing reusable Nix modules, debugging distributed systems, or making sure a platform can be rebuilt predictably from scratch.

Critical components of the environment may include high-performance compute, distributed Linux file systems, network design, security-in-depth, ML/AI infrastructure, high-bandwidth data processing, cloud deployment, and fielded systems.

You don't need to be an expert at everything. Whatever part you bite off, though, will be yours to own. This is a great opportunity to apply deep systems expertise while learning adjacent areas.

Our team needs your help making sure our customers are repeatedly provided with operational insights from a multi-domain data environment-and that the infrastructure producing those insights is dependable, observable, reproducible, and recoverable.

These Are the Things You' ll Get to Do:
  • Own the reliability and operation of critical compute and data platforms.
  • Design, deploy, and maintain Linux infrastructure, including NixOS-based systems.
  • Build reproducible system configurations and deployment workflows using Nix and infrastructure-as-code.
  • Develop reusable NixOS modules, packages, flakes, and system configurations.
  • Improve platform resilience, fault tolerance, recoverability, and maintainability.
  • Build monitoring, logging, alerting, and observability capabilities that help us understand system behavior before customers notice problems.
  • Diagnose difficult failures across Linux, networking, storage, containers, and distributed systems.
  • Automate repetitive operational tasks and eliminate configuration drift.
  • Design and build security and maintainability components for deployed systems.
  • Help define operational standards, deployment practices, upgrade strategies, and disaster-recovery procedures.
  • Perform capacity planning and identify performance bottlenecks across compute, storage, and networking.
  • Own solutions spanning CNO, analytics, security, infrastructure, and field deployments.
  • Learn new skills and teach the rest of us what you know.
Minimum Qualifications:
  • Enthusiasm for learning new stuff.
  • Bachelor's degree in Computer Science, Computer Engineering, a related field, or amazing equivalent skills.
  • Strong experience administering and troubleshooting Linux systems.
  • Experience with Linux networking, storage, and file systems.
  • Experience automating system configuration and deployment.
  • Experience with Nix and/or NixOS in production, lab, or significant personal environments.
  • Python, Go, Bash, or similar scripting/programming experience of 2+ years.
  • Ability to debug complex systems methodically across multiple layers of the stack.
Nice-to-Have Qualifications:
  • Deep experience with NixOS, including custom modules, overlays, flakes, packaging, and reproducible deployments.
  • Experience operating fleets of Linux systems.
  • Infrastructure-as-code experience with Nix, Terraform, Ansible, or similar tooling.
  • Experience designing highly available or fault-tolerant systems.
  • Monitoring and observability experience with tools such as Prometheus, Grafana, Loki, OpenTelemetry, Elasticsearch, or similar systems.

If you're looking for comfort, keep scrolling. At Leidos, we outthink, outbuild, and outpace the status quo - because the mission demands it. We're not hiring followers. We're recruiting the ones who disrupt, provoke, and refuse to fail. Step 10 is ancient history. We're already at step 30 - and moving faster than anyone else dares.

Original Posting:

August 14, 2026

For U.S. Positions: While subject to change based on business needs, Leidos reasonably anticipates that this job requisition will remain open for at least 3 days with an anticipated close date of no earlier than 3 days after the original posting date as listed above.

Pay Range:

Pay Range $87,100.00 - $157,450.00

About Leidos

Leidos is an industry and technology leader serving government and commercial customers with smarter, more efficient digital and mission innovations. Headquartered in Reston, Virginia, with 47,000 global employees, Leidos reported annual revenues of approximately $16.7 billion for the fiscal year ended January 3, 2025. For more information, visit www.Leidos.com.

Pay and Benefits

Pay and benefits are fundamental to any career decision. That's why we craft compensation packages that reflect the importance of the work we do for our customers. Employment benefits include competitive compensation, Health and Wellness programs, Income Protection, Paid Leave and Retirement. More details are available at www.leidos.com/careers/pay-benefits.

Securing Your Data

Beware of fake employment opportunities using Leidos' name. Leidos will never ask you to provide payment-related information during any part of the employment application process (i.e., ask you for money), nor will Leidos ever advance money as part of the hiring process (i.e., send you a check or money order before doing any work). Further, Leidos will only communicate with you through emails that are generated by the Leidos.com automated system - never from free commercial services (e.g., Gmail, Yahoo, Hotmail) or via WhatsApp, Telegram, etc. If you received an email purporting to be from Leidos that asks for payment-related information or any other personal information (e.g., about you or your previous employer), and you are concerned about its legitimacy, please make us aware immediately by emailing us at LeidosCareersFraud@leidos.com.

If you believe you are the victim of a scam, contact your local law enforcement and report the incident to the U.S. Federal Trade Commission.

Commitment to Non-Discrimination

All qualified applicants will receive consideration for employment without regard to sex, race, ethnicity, age, national origin, citizenship, religion, physical or mental disability, medical condition, genetic information, pregnancy, family structure, marital status, ancestry, domestic partner status, sexual orientation, gender identity or expression, veteran or military status, or any other basis prohibited by law. Leidos will also consider for employment qualified applicants with criminal histories consistent with relevant laws.

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.
oder ziehe deine Datei hierhin.
Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

AI & Security Infrastructure Integration Engineer
AI & Security Infrastructure Integration Engineer

Leidos • Deutschland

Hybrid
EUR 93.000 - 169.000
ServiceNow Administrator - Senior
ServiceNow Administrator - Senior

Leidos • Stuttgart

Vor Ort
EUR 59.000 - 109.000
Competitive compensation
Health and Wellness programs
Retirement benefits
Cloud/SecDevOps Engineer
Cloud/SecDevOps Engineer

Leidos • Deutschland

Hybrid
EUR 60.000 - 111.000
Information Systems Security Officer
Information Systems Security Officer

Leidos • Deutschland

Hybrid
EUR 60.000 - 109.000
ASI Manager
ASI Manager

Leidos • Stuttgart

Vor Ort
EUR 51.000 - 92.000
Defensive Cyber Operations Analyst
Defensive Cyber Operations Analyst

Leidos • Deutschland

Hybrid
EUR 75.000 - 136.000
Robotics Computer Engineer
Robotics Computer Engineer

Leidos • Deutschland

Remote
EUR 93.000 - 170.000
Competitive compensation
Health and Wellness programs
Income Protection
+1
Principal Network Engineer
Principal Network Engineer

Leidos • Deutschland

Hybrid
EUR 120.000 - 217.000
Test and Evaluation Lead
Test and Evaluation Lead

Leidos • Deutschland

Hybrid
EUR 101.000 - 182.000
TRICARE Beneficiary Services Representative - Grafenwoehr, Germany
TRICARE Beneficiary Services Representative - Grafenwoehr, Germany

Leidos • Grafenwöhr

Vor Ort
EUR 28.000 - 50.000
Foreign service premium