ME00680-Site Reliability Engineer 3

Momentumcareers

Corridor North (MD)

On-site

USD 150,000 - 190,000

Full time

6 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

11 paid holidays
3 weeks PTO
Medical, dental, vision plans

Job summary

Momentum Engineering, Inc. seeks a Site Reliability Engineer to support a cloud-based platform built with Java/FOSS tech. You will enable data-intensive analytics across managed infrastructure in a mission-focused environment.

The ideal candidate thrives in a fast-paced team, is self‑motivated, and proactively resolves operational issues. This on-call role includes Tier 1–3 support responsibilities and cross-disciplinary collaboration.

Qualifications

  • Active Top Secret/SCI clearance with NSA Full Scope Polygraph required.
  • Bachelor’s degree in Computer Science or related field; equivalent experience may be considered.
  • Master’s degree may be equivalent to four years of experience.
  • DoD 8570 IAT Level I certification or higher preferred.
  • Fourteen (14) years of relevant technical experience.
  • Strong experience troubleshooting operational issues in Linux environments.

Responsibilities

  • Operate, administer, and ensure reliability of cloud-based infrastructure and platform services.
  • Provide Tier 1–3 technical support for users, applications, and infrastructure.
  • Troubleshoot complex system, application, networking, and Linux issues.
  • Monitor system health, availability, and performance; respond to incidents.
  • Support Kubernetes, Hadoop, and Accumulo in distributed environments.
  • Manage containerized apps with Docker and Kubernetes.
  • Develop automation scripts using Python, Bash, or similar.
  • Support HDFS environments and use Prometheus/Grafana for observability.
  • Use Salt and Ansible for configuration management and automation.
  • Contribute to incident response, root cause analysis, and improvements.

Skills

Top Secret/SCI clearance
Kubernetes
Linux troubleshooting
Python scripting
Bash scripting
Tier 1–3 support
Incident response
Automation
Cloud platforms

Education

Bachelor’s degree in Computer Science or related field
Master’s degree may be equivalent to four years experience
DoD 8570 IAT Level I certification or higher

Tools

Docker
Kubernetes
Hadoop
Accumulo
HDFS
Prometheus
Grafana
Salt
Ansible
OpenStack
AWS

Job description

Momentum Engineering, Inc. fosters an employee-centric culture. Our strength lies in our people. With a high percentage of employees holding advanced degrees in engineering, computer science, and related disciplines, we bring deep technical expertise to every mission. Our team includes professionals with security clearances and full-scope polygraphs, ensuring trusted, secure support for the most sensitive national security initiatives. Additionally, our workforce is equipped with industry-leading certifications, demonstrating a commitment to continuous learning and excellence. Most importantly, our exceptional employee retention rate reflects a culture of professional growth, mission focus, and dedication—ensuring long-term stability and expertise for our customers’ critical needs.

Job Summary
  • Seeking a Site Reliability Engineer to support a cloud-based platform built with Java and Free and Open-Source Software technologies, including Kubernetes, Hadoop, and Accumulo
  • The platform enables the execution of data-intensive analytics across managed infrastructure in a mission-focused environment
  • The ideal candidate is self-motivated, detail-oriented, and able to thrive in a fast-paced team environment while proactively identifying and resolving operational issues
  • This is an on-call position and includes Tier 1 through Tier 3 support responsibilities
Primary Responsibilities
  • Support the operation, administration, reliability, and availability of cloud-based infrastructure and platform services
  • Provide Tier 1 through Tier 3 technical support for operational issues affecting users, applications, and infrastructure
  • Troubleshoot and resolve complex system, application, networking, and infrastructure issues within Linux environments
  • Monitor system health, availability, performance, and operational status
  • Support distributed computing technologies including Kubernetes, Hadoop, and Accumulo
  • Support containerized applications and services using Docker and Kubernetes
  • Develop and maintain automation and administrative scripts using Python, Bash, or similar scripting languages
  • Support Hadoop Distributed File System (HDFS) environments
  • Use monitoring and observability tools such as Prometheus and Grafana to identify and resolve operational issues
  • Support configuration management and automation tools such as Salt and Ansible
  • Participate in incident response, root cause analysis, corrective actions, and continuous improvement activities
  • Support virtualization, cloud, and hybrid infrastructure environments
  • Document troubleshooting procedures, operational processes, system changes, and recurring issues
  • Collaborate with developers, system administrators, engineers, and mission stakeholders to maintain reliable and scalable platform operations
  • Participate in on-call support and respond to operational issues as required
Required Qualifications
  • Must have active Top Secret/SCI clearance with NSA Full Scope Polygraph
  • A Bachelor’s degree in Computer Science or a related technical field is highly desired and may be considered equivalent to two (2) years of experience
  • A Master’s degree in a technical field may be considered equivalent to four (4) years of experience
  • Degrees in Mathematics, Information Systems, Engineering, or similar disciplines will be considered technical degrees
  • Fourteen (14) years of relevant technical experience
  • Strong experience troubleshooting operational issues in Linux environments
  • DoD 8570 IAT Level I certification or higher
  • Ability to provide Tier 1 through Tier 3 support in a mission-critical environment
  • Candidates must possess at least one of the following certifications:
    • AWS Certified Developer – Associate
    • AWS Certified Solutions Architect – Associate
    • AWS Certified Solutions Architect – Professional
    • AWS Certified SysOps Administrator – Associate
    • Certified Kubernetes Application Developer (CKAD)
    • Elastic Certified Engineer
    • Elastic Certified Observability Engineer
Desired Qualifications
  • Experience with one or more of the following technologies is beneficial:
    • Docker
    • Kubernetes
    • Hadoop
    • Apache Accumulo
    • Hadoop Distributed File System (HDFS)
    • Python
    • Bash
    • Prometheus
    • Grafana
    • JIRA
    • Salt
    • Ansible
    • Virtualization technologies
    • OpenStack
    • Amazon Web Services (AWS)

Exempt hourly position. 11 paid holidays, minimum of 3 weeks PTO, company sponsored group medical plan, company paid dental, vision, life insurance, and STD/LTD plans. Salary is dependent upon the candidate’s experience and qualifications.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

ME00680-Site Reliability Engineer 3
ME00680-Site Reliability Engineer 3

Momentum Engineering • Maryland

On-site
USD 165,000 - 230,000
11 paid holidays
Minimum 3 weeks PTO
Group medical plan
+1
ME00616-Cloud System Administrator 2
ME00616-Cloud System Administrator 2

Momentum Engineering, Inc. • Maryland

On-site
USD 90,000 - 120,000
11 paid holidays
3 weeks PTO
Company sponsored medical plan
+3
ME00616-Cloud System Administrator 2
ME00616-Cloud System Administrator 2

Momentum Engineering, Inc • Annapolis (MD)

On-site
USD 85,000 - 110,000
11 paid holidays
Minimum of 3 weeks PTO
Company sponsored medical plan
+1
ME00678-Cloud System Administrator 2
ME00678-Cloud System Administrator 2

Momentumcareers • Corridor North (MD)

On-site
USD 120,000 - 160,000
11 paid holidays
3 weeks PTO
Group medical plan
+4
ME00617-Cloud System Administrator 2
ME00617-Cloud System Administrator 2

Momentum Engineering, Inc. • Maryland

On-site
USD 80,000 - 110,000
11 paid holidays
Minimum of 3 weeks PTO
Company sponsored group medical plan
+1
ME00616-Cloud System Administrator 2
ME00616-Cloud System Administrator 2

Momentum Engineering • Maryland

On-site
USD 150,000 - 205,000
11 paid holidays
3 weeks PTO
Company sponsored medical plan
+2
ME00678-Cloud System Administrator 2
ME00678-Cloud System Administrator 2

Momentum Engineering • Corridor North (MD)

On-site
USD 150,000 - 205,000
11 paid holidays
3 weeks PTO
Company medical plan
+4
ME00614-Cloud Software Engineer 3
ME00614-Cloud Software Engineer 3

Momentum Engineering, Inc. • Maryland

On-site
USD 110,000 - 160,000
11 paid holidays
Minimum of 3 weeks PTO
Company sponsored medical plan
+2
ME00614-Cloud Software Engineer 3
ME00614-Cloud Software Engineer 3

Momentum Engineering, Inc • Annapolis (MD)

On-site
USD 110,000 - 150,000
11 paid holidays
Minimum of 3 weeks PTO
Company sponsored medical plan
+2
ME00617-Cloud System Administrator 2
ME00617-Cloud System Administrator 2

Momentum Engineering, Inc • Annapolis (MD)

On-site
USD 85,000 - 120,000
11 paid holidays
Minimum of 3 weeks PTO
Company sponsored medical plan
+2