ME00680-Site Reliability Engineer 3

Momentum Engineering

Maryland

On-site

USD 165,000 - 230,000

Full time

5 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

11 paid holidays
Minimum 3 weeks PTO
Group medical plan
Dental, vision, life insurance

Job summary

Momentum Engineering seeks a Site Reliability Engineer 3 to support a cloud-based analytics platform built with Java, Kubernetes, Hadoop, and Accumulo. The role emphasizes proactive issue prevention, automation, and on-call support in a mission-focused environment.

The candidate will work with Linux, Docker, Kubernetes, and OpenStack-related technologies, ensuring reliability and scalability of critical systems for national security initiatives.

Qualifications

  • Must have active TS/SCI clearance with NSA FSP
  • Bachelor’s in CS or related field; advanced degree or 14+ years experience considered equivalent
  • DoD 8570 IAT Level I or higher
  • Experience troubleshooting Linux environments and incident response
  • Certifications: CKAD, AWS... etc (list from ad)

Responsibilities

  • Provide Tier 1–3 support for cloud-based infrastructure
  • Troubleshoot complex Linux, network, and app issues
  • Monitor health, performance, and availability of platform
  • Develop automation scripts in Python/Bash
  • Maintain docs for procedures and incidents
  • Collaborate with engineers and stakeholders
  • Participate in on-call duty

Skills

Top Secret/SCI clearance
DoD 8570 IAT Level I
Incident response
Hybrid/mission-critical mindset

Education

Bachelor’s degree in Computer Science or related field
Master’s degree may be considered equivalent
14+ years of relevant technical experience

Tools

Kubernetes
Docker
Hadoop
Apache Accumulo
Python
Bash
Prometheus
Grafana
Salt
Ansible
OpenStack
AWS
HDFS

Job description

ME00680-Site Reliability Engineer 3

Momentum Engineering Annapolis Junction, Maryland, United States

About this position

Momentum Engineering, Inc. fosters an employee-centric culture. Our strength lies in our people. With a high percentage of employees holding advanced degrees in engineering, computer science, and related disciplines, we bring deep technical expertise to every mission. Our team includes professionals with security clearances and full-scope polygraphs, ensuring trusted, secure support for the most sensitive national security initiatives. Additionally, our workforce is equipped with industry-leading certifications, demonstrating a commitment to continuous learning and excellence. Most importantly, our exceptional employee retention rate reflects a culture of professional growth, mission focus, and dedication—ensuring long‑term stability and expertise for our customers’ critical needs.

Job Summary
  • Seeking a Site Reliability Engineer to support a cloud‑based platform built with Java and Free and Open‑Source Software technologies, including Kubernetes, Hadoop, and Accumulo
  • The platform enables the execution of data‑intensive analytics across managed infrastructure in a mission‑focused environment
  • The ideal candidate is self‑motivated, detail‑oriented, and able to thrive in a fast‑paced team environment while proactively identifying and resolving operational issues
  • This is an on‑call position and includes Tier 1 through Tier 3 support responsibilities
Primary Responsibilities
  • Support the operation, administration, reliability, and availability of cloud‑based infrastructure and platform services
  • Provide Tier 1 through Tier 3 technical support for operational issues affecting users, applications, and infrastructure
  • Troubleshoot and resolve complex system, application, networking, and infrastructure issues within Linux environments
  • Monitor system health, availability, performance, and operational status
  • Support distributed computing technologies including Kubernetes, Hadoop, and Accumulo
  • Support containerized applications and services using Docker and Kubernetes
  • Develop and maintain automation and administrative scripts using Python, Bash, or similar scripting languages
  • Use monitoring and observability tools such as Prometheus and Grafana to identify and resolve operational issues
  • Support configuration management and automation tools such as Salt and Ansible
  • Participate in incident response, root cause analysis, corrective actions, and continuous improvement activities
  • Support virtualization, cloud, and hybrid infrastructure environments
  • Document troubleshooting procedures, operational processes, system changes, and recurring issues
  • Collaborate with developers, system administrators, engineers, and mission stakeholders to maintain reliable and scalable platform operations
  • Participate in on‑call support and respond to operational issues as required
Required Qualifications
  • Must have active Top Secret/SCI clearance with NSA Full Scope Polygraph
  • A Bachelor’s degree in Computer Science or a related technical field is highly desired and may be considered equivalent to two (2) years of experience
  • A Master’s degree in a technical field may be considered equivalent to four (4) years of experience
  • Degrees in Mathematics, Information Systems, Engineering, or similar disciplines will be considered technical degrees
  • Fourteen (14) years of relevant technical experience
  • Strong experience troubleshooting operational issues in Linux environments
  • DoD 8570 IAT Level I certification or higher
  • Ability to provide Tier 1 through Tier 3 support in a mission‑critical environment
  • Candidates must possess at least one of the following certifications:
    • AWS Certified Developer – Associate
    • AWS Certified Solutions Architect – Associate
    • AWS Certified Solutions Architect – Professional
    • AWS Certified SysOps Administrator – Associate
    • Certified Kubernetes Application Developer (CKAD)
    • Elastic Certified Engineer
    • Elastic Certified Observability Engineer
Desired Qualifications
  • Experience with one or more of the following technologies is beneficial:
    • Docker
    • Kubernetes
    • Hadoop
    • Apache Accumulo
    • Hadoop Distributed File System (HDFS)
    • Python
    • Bash
    • Prometheus
    • Grafana
    • JIRA
    • Salt
    • Ansible
    • Virtualization technologies
    • OpenStack
    • Amazon Web Services (AWS)

Exempt hourly position. 11 paid holidays, minimum of 3 weeks PTO, company sponsored group medical plan, company paid dental, vision, life insurance, and STD/LTD plans. Salary is dependent upon the candidate’s experience and qualifications.

Momentum Engineering, Inc. fosters an employee‑centric culture. Our strength lies in our people. With a high percentage of employees holding advanced degrees in engineering, computer science, and related disciplines, we bring deep technical expertise to every mission. Our team includes professionals with security clearances and full‑scope polygraphs, ensuring trusted, secure support for the most sensitive national security initiatives. Additionally, our workforce is equipped with industry‑leading certifications, demonstrating a commitment to continuous learning and excellence. Most importantly, our exceptional employee retention rate reflects a culture of professional growth, mission focus, and dedication—ensuring long‑term stability and expertise for our customers’ critical needs.

Job Summary
  • Seeking a Site Reliability Engineer to support a cloud‑based platform built with Java and Free and Open‑Source Software technologies, including Kubernetes, Hadoop, and Accumulo
  • The platform enables the execution of data‑intensive analytics across managed infrastructure in a mission‑focused environment
  • The ideal candidate is self‑motivated, detail‑oriented, and able to thrive in a fast‑paced team environment while proactively identifying and resolving operational issues
  • This is an on‑call position and includes Tier 1 through Tier 3 support responsibilities
Primary Responsibilities
  • Support the operation, administration, reliability, and availability of cloud‑based infrastructure and platform services
  • Provide Tier 1 through Tier 3 technical support for operational issues affecting users, applications, and infrastructure
  • Troubleshoot and resolve complex system, application, networking, and infrastructure issues within Linux environments
  • Monitor system health, availability, performance, and operational status
  • Support distributed computing technologies including Kubernetes, Hadoop, and Accumulo
  • Support containerized applications and services using Docker and Kubernetes
  • Develop and maintain automation and administrative scripts using Python, Bash, or similar scripting languages
  • Support Hadoop Distributed File System (HDFS) environments
  • Use monitoring and observability tools such as Prometheus and Grafana to identify and resolve operational issues
  • Support configuration management and automation tools such as Salt and Ansible
  • Participate in incident response, root cause analysis, corrective actions, and continuous improvement activities
  • Support virtualization, cloud, and hybrid infrastructure environments
  • Document troubleshooting procedures, operational processes, system changes, and recurring issues
  • Collaborate with developers, system administrators, engineers, and mission stakeholders to maintain reliable and scalable platform operations
  • Participate in on‑call support and respond to operational issues as required
Required Qualifications
  • Must have active Top Secret/SCI clearance with NSA Full Scope Polygraph
  • A Bachelor’s degree in Computer Science or a related technical field is highly desired and may be considered equivalent to two (2) years of experience
  • A Master’s degree in a technical field may be considered equivalent to four (4) years of experience
  • Degrees in Mathematics, Information Systems, Engineering, or similar disciplines will be considered technical degrees
  • Fourteen (14) years of relevant technical experience
  • Strong experience troubleshooting operational issues in Linux environments
  • DoD 8570 IAT Level I certification or higher
  • Ability to provide Tier 1 through Tier 3 support in a mission‑critical environment
  • Candidates must possess at least one of the following certifications:
    • AWS Certified Developer – Associate
    • AWS Certified Solutions Architect – Associate
    • AWS Certified Solutions Architect – Professional
    • AWS Certified SysOps Administrator – Associate
    • Certified Kubernetes Application Developer (CKAD)
    • Elastic Certified Engineer
    • Elastic Certified Observability Engineer
Desired Qualifications
  • Experience with one or more of the following technologies is beneficial:
    • Docker
    • Kubernetes
    • Hadoop
    • Apache Accumulo
    • Hadoop Distributed File System (HDFS)
    • Python
    • Bash
    • Prometheus
    • Grafana
    • JIRA
    • Salt
    • Ansible
    • Virtualization technologies
    • OpenStack
    • Amazon Web Services (AWS)

Exempt hourly position. 11 paid holidays, minimum of 3 weeks PTO, company sponsored group medical plan, company paid dental, vision, life insurance, and STD/LTD plans. Salary is dependent upon the candidate’s experience and qualifications.

The pay range for this role is:
165,000 - 230,000 USD per year(NBP)

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

ME00678-Cloud System Administrator 2
ME00678-Cloud System Administrator 2

Momentum Engineering • Corridor North (MD)

On-site
USD 150,000 - 205,000
11 paid holidays
3 weeks PTO
Company medical plan
+4
ME00680-Site Reliability Engineer 3
ME00680-Site Reliability Engineer 3

Momentumcareers • Corridor North (MD)

On-site
USD 150,000 - 190,000
11 paid holidays
3 weeks PTO
Medical, dental, vision plans
ME00616-Cloud System Administrator 2
ME00616-Cloud System Administrator 2

Momentum Engineering • Maryland

On-site
USD 150,000 - 205,000
11 paid holidays
3 weeks PTO
Company sponsored medical plan
+2
ME00617-Cloud System Administrator 2
ME00617-Cloud System Administrator 2

Momentum Engineering • Maryland

On-site
USD 150,000 - 205,000
11 paid holidays
Minimum of 3 weeks PTO
Company sponsored medical plan
+1
ME00678-Cloud System Administrator 2
ME00678-Cloud System Administrator 2

Momentumcareers • Corridor North (MD)

On-site
USD 120,000 - 160,000
11 paid holidays
3 weeks PTO
Group medical plan
+4
ME00616-Cloud System Administrator 2
ME00616-Cloud System Administrator 2

Momentum Engineering, Inc • Annapolis (MD)

On-site
USD 85,000 - 110,000
11 paid holidays
Minimum of 3 weeks PTO
Company sponsored medical plan
+1
ME00616-Cloud System Administrator 2
ME00616-Cloud System Administrator 2

Momentum Engineering, Inc. • Maryland

On-site
USD 90,000 - 120,000
11 paid holidays
3 weeks PTO
Company sponsored medical plan
+3
ME00614-Cloud Software Engineer 3
ME00614-Cloud Software Engineer 3

Momentum Engineering • Maryland

On-site
USD 180,000 - 235,000
11 paid holidays
Minimum of 3 weeks paid time off (PTO)
Company-sponsored group medical plan
+1
ME00617-Cloud System Administrator 2
ME00617-Cloud System Administrator 2

Momentum Engineering, Inc • Annapolis (MD)

On-site
USD 85,000 - 120,000
11 paid holidays
Minimum of 3 weeks PTO
Company sponsored medical plan
+2
ME00615-Cloud System Administrator 2
ME00615-Cloud System Administrator 2

Momentum Engineering • Maryland

On-site
USD 150,000 - 205,000
11 paid holidays
Minimum of 3 weeks PTO
Company-sponsored medical plan
+2