Senior Site Reliability Engineer – Secret Clearance Required

Jobtailor

Arlington (VA)

On-site

USD 140,000 - 190,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Jobtailor in Arlington, VA is seeking a Site Reliability Engineer focused on reliability, incident response, and automation. You will work directly in code (TypeScript) to improve performance and reliability across the stack.

Responsibilities include building observability with Prometheus, Loki, Alloy, and Grafana, defining SLIs/SLOs, and leading blameless post-mortems. Collaboration across product, platform, and DevOps teams is essential.

Qualifications

  • Extensive experience with TypeScript and real-time shipping of application code.
  • 5+ years in software engineering, SRE, or related role.
  • Strong SDLC understanding across design, code review, testing, and release.
  • Experience with incident response, root cause analysis, and lasting fixes.
  • Collaborator who works well across product, platform, and DevOps teams.

Responsibilities

  • Improve the application by fixing reliability and performance at the source.
  • Build observability with monitoring, logging, and alerting (Prometheus, Loki, Alloy, Grafana).
  • Define and measure SLIs and SLOs and set reliable targets.
  • Lead incident response and blameless post-mortems.
  • Automate repetitive operational work to reduce toil.

Skills

TypeScript Programming
Incident Response
Reliability Engineering
Monitoring & Logging
SDLC Expertise

Education

Active Secret Clearance

Tools

Prometheus
Loki
Alloy
Grafana

Job description

  • Improving the application: Work directly in the codebase (primarily TypeScript) to fix reliability and performance problems at the source.
  • Building observability that developers actually use: Design and run our monitoring, logging, and alerting (Prometheus, Loki, Alloy, Grafana).
  • Owning reliability targets: Define and measure SLIs and SLOs, wire up alerting that feeds them, and be the person who can say what 'reliable' means for our systems.
  • Leading incident response: Act as incident responder, and incident commander when needed. Run blameless post-mortems (AARs).
  • Automating away toil: Spot the repetitive operational work and write software to kill it.
Requirements
  • An active Secret clearance
  • 5+ years in software engineering, SRE, or a related role, with real time spent writing and shipping application code
  • Strong TypeScript (or comparable modern language experience with willingness to work primarily in TypeScript)
  • Solid grasp of the full SDLC: design, code review, testing, release, and how reliability fits into each stage
  • Experience with incident response, root cause analysis, and turning findings into lasting fixes
  • A collaborator who works well across product, platform, and DevOps teams and shares context openly

Demonstrates expertise in TypeScript and the full Software Development Life Cycle (SDLC), with a strong focus on reliability, incident response, and automation. Proven ability to collaborate effectively across teams while owning reliability targets and improving application performance.

Highest-signal resume keywords
  • TypeScript Programming
  • Incident Response
  • Reliability Engineering
  • Monitoring and Logging
  • Software Development Life Cycle (SDLC)
ATS Optimization Keywords
Hard Skills
  • TypeScript
  • Software Engineering
  • Incident Response
  • Root Cause Analysis
  • Performance Optimization
  • Application Reliability
  • Automation
  • Code Review
  • Testing
  • Release Management
Soft Skills
  • Collaboration
  • Communication
  • Context Sharing
Certifications & Qualifications
  • Active Secret Clearance
Industry Keywords
  • Site Reliability Engineering (SRE)
  • Service Level Indicators (SLIs)
  • Service Level Objectives (SLOs)
  • Blameless Post-Mortems
  • Application Monitoring
Tools & Technologies
  • Prometheus
  • Loki
  • Alloy
  • Grafana
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer – Lead
Site Reliability Engineer – Lead

Jobtailor • Arizona

On-site
USD 140,000 - 230,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Jobtailor • Arlington (VA)

On-site
USD 140,000 - 200,000
Staff Site Reliability Engineer
Staff Site Reliability Engineer

Jobtailor • California (MO)

On-site
USD 140,000 - 210,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Jobtailor • Town of Florida (NY)

Hybrid
USD 150,000 - 190,000
Site Reliability Engineer
Site Reliability Engineer

TalentDome Staffing • United States

On-site
USD 140,000 - 210,000
Site Reliability Engineer (Secret Clearance)
Site Reliability Engineer (Secret Clearance)

ROI Services LLC • Huntsville (AL)

On-site
USD 110,000 - 150,000
Director, Site Reliability Engineering
Director, Site Reliability Engineering

Jobtailor • California (MO)

On-site
USD 180,000 - 260,000
Site Reliability Engineer
Site Reliability Engineer

Jobtailor • New Hampshire

On-site
USD 110,000 - 160,000
Lead Site Reliability Engineer
Lead Site Reliability Engineer

Luxoft • Buffalo (NY)

On-site
USD 140,000 - 190,000
Lead Site Reliability Engineer
Lead Site Reliability Engineer

Luxoft • Wilmington (DE)

On-site
USD 140,000 - 190,000