Site Reliability Engineer (SRE) - SecOps

Arkenstone Defense

Albuquerque (NM)

On-site

USD 120,000 - 180,000

Full time

2 hours ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Competitive Salary
Health & Wellness
401(k) Matching
Paid Time Off
Employee Assistance
Professional Development

Job summary

Arkenstone Defense is seeking a highly motivated Site Reliability Engineer (SRE) to lead the design and implementation of operational excellence across AWS, Azure, and GCP. You will own uptime, observability, and resilience for critical services, partnering with product owners, developers, and security teams.

You will drive automation, define SLOs/SLAs, and champion IaC practices while mentoring junior engineers and shaping a culture of accountability and continuous improvement.

Qualifications

  • 5+ years in SRE, DevOps, or infra roles.
  • Hands-on multi-cloud experience (AWS & GCP) and Kubernetes.
  • Strong scripting (Python, Bash) and IaC (Terraform).
  • Experience with monitoring stacks (Prometheus, Grafana, Datadog, ELK).

Responsibilities

  • Design and own multi-cloud infra reliability across AWS, Azure, GCP.
  • Develop logging, monitoring, and alerting systems.
  • Define SLOs/SLAs for critical services.
  • Lead incident response and postmortems.
  • Automate deployment, scaling, and recovery workflows.
  • Mentor engineers and promote IaC practices.

Skills

AWS
GCP
Kubernetes
Terraform
Python
CI/CD
Observability
Incident response

Tools

Prometheus
Grafana
Datadog
ELK
CloudFormation

Job description

At Arkenstone Defense, we empower defense tech startups with the tools, infrastructure, and compliance solutions they need to become successful prime contractors. Our mission is to remove barriers and help innovators grow - from day one to becoming a trusted prime for the U.S. Government.

We're early, we're lean, and we're building something that actually matters. The people who do well here aren't waiting to be told what to do; they see a gap and fill it.

Overview
About Us

At Arkenstone Defense, we empower defense tech startups with the tools, infrastructure, and compliance solutions they need to become successful prime contractors. Our mission is to remove barriers and help innovators grow - from day one to becoming a trusted prime for the U.S. Government.

We're early, we're lean, and we're building something that actually matters. The people who do well here aren't waiting to be told what to do; they see a gap and fill it.

Overview

We are seeking a highly motivated Systems Reliability Engineer (SRE) to lead the design and implementation of operational excellence across our multi-cloud environments. This role is central to ensuring the scalability, reliability, and performance of our products running in AWS, Azure, and GCP infrastructure.

As the lead SRE, you will own the uptime, observability, and system resilience for our critical services. This includes driving architecture decisions, automation practices, and incident response strategies - working closely with the product owner(s), developer teams, and security operations teams.

What You’ll Do
  • Design, implement, and own the infrastructure reliability strategy across AWS, Azure, and GCP
  • Champion observability by developing and maintaining effective logging, monitoring, and alerting systems
  • Define and enforce SLOs/SLAs for critical systems and services
  • Lead efforts in performance tuning, system hardening, capacity planning, and disaster recovery
  • Own the incident management lifecycle: from detection to postmortem and root cause analysis
  • Automate deployment, scaling, and recovery workflows to reduce manual toil
  • Contribute to infrastructure as code (Terraform, ARM templates, CloudFormation, etc.)
  • Act as a mentor and technical leader to junior engineers and cross-functional partners
  • Drive a culture of accountability, ownership, and continuous improvement
  • Perform any other related duties as required or assigned.
Requirements
  • 5+ years of experience in SRE, DevOps, or infrastructure engineering roles
  • Proven track record of operating large-scale systems in multi-cloud environments, with hands-on expertise in AWS and GCP
  • Strong knowledge of cloud-native architecture, container orchestration with Kubernetes, and CI/CD pipelines
  • Proficient in scripting (Python, Bash, etc.) and infrastructure automation tools (e.g., Terraform)
  • Experience with monitoring and observability platforms (e.g., Prometheus, Grafana, Datadog, ELK)
  • Excellent problem-solving skills with the ability to manage incidents and make sound decisions under pressure
  • Clear communicator capable of translating technical concepts to mixed audiences and participating in customer discussions
Who You Are
  • The Security Builder: You don't just consume security tools — you extend and improve them. You're energized by the opportunity to make analysts faster and compliance more automated.
  • Quality Over Speed: You write code that lasts. You push for clean interfaces, good documentation, and tests — even in a fast-moving environment.
  • Security-Minded Developer: You treat security as a first-class requirement, not an afterthought. You're comfortable reading CVEs, threat models, and compliance controls.
  • Cross-Functional Partner: You can work fluidly with security analysts, engineers, and compliance professionals, translating needs into reliable software.
Mission Alignment

We are a Defense-focused company supporting sensitive and cleared workforces. The Site Reliability Engineer (SRE) - SecOps will embrace our commitment to operational excellence, compliance rigor, and a world-class employee experience.

Physical Requirements
  • Prolonged periods of sitting at a desk and working on a computer
  • Must be able to lift up to 15 pounds at times
  • May require occasional travel to office locations or client sites
  • Ability to communicate effectively in written and verbal form
Benefits for working with us!

We are committed to supporting our employees both professionally and personally. Our robust benefits package is designed to promote your well-being, growth, and work-life balance

  • Competitive Salary: Recognizing your hard work with attractive compensation and rewarding excellence.
  • Health and Wellness Programs: Including medical, dental, & vision insurance options, along with mental health support & wellness initiatives.
  • Retirement Planning: Secure your future with our flexible 401(k) plan and matching company contributions.
  • Paid Time Off & Holidays: Generous PTO, sick leave, and holiday pay to help you recharge and enjoy life outside of work.
  • Employee Assistance Program: Confidential resources for personal and professional support.
  • Professional Development: Access to training, certifications, and continuing education to foster your career growth.

We are an Equal Opportunity Employer. We celebrate diversity and are committed to creating an inclusive environment for all employees. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex (including pregnancy, gender identity, and sexual orientation), national origin, age, disability, genetic information, veteran status, or any other characteristic protected under applicable law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer SRE SecOps
Site Reliability Engineer SRE SecOps

Arkenstone • Menlo Park (CA)

On-site
USD 140,000 - 190,000
Competitive Salary
Health & Wellness Programs
401(k) Plan
+3
Staff Software Engineer (SWE) - SecOps
Staff Software Engineer (SWE) - SecOps

Arkenstone Defense • Menlo Park (CA)

On-site
USD 150,000 - 210,000
Competitive Salary
Health and Wellness Programs
Retirement Planning
+3
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

OutSolve • Mission (KS)

Remote
USD 90,000 - 130,000
100% remote work environment
Competitive compensation
Professional development opportunities
+1
Site Reliability Engineer
Site Reliability Engineer

VantageScore® • San Francisco (CA)

On-site
USD 150,000
Medical insurance
Dental insurance
401(k) plan
+1
Senior DevOps/SRE Engineer
Senior DevOps/SRE Engineer

SEI • Chicago (IL)

Hybrid
USD 140,000 - 170,000
Comprehensive healthcare benefits
401(k) match
Paid Time Off (PTO)
+2
Senior SRE & SecOps Architect - Multi-Cloud Reliability
Senior SRE & SecOps Architect - Multi-Cloud Reliability

Arkenstone • Menlo Park (CA)

On-site
USD 140,000 - 190,000
Competitive Salary
Health & Wellness Programs
401(k) Plan
+3
Site Reliability Engineer
Site Reliability Engineer

VantageScore • San Francisco (CA)

On-site
USD 135,000 - 165,000
Medical benefits
401(k)
Paid time off
Senior DevOps/SRE Engineer
Senior DevOps/SRE Engineer

VITG • Ellicott City (MD)

Hybrid
USD 90,000 - 120,000
401(k) with employer contribution
Medical/Dental/Vision insurance
Paid vacation (PTO)
Site Reliability Engineering (SRE) Architect
Site Reliability Engineering (SRE) Architect

Cloud Analytics Technologies, LLC • Atlanta (GA)

On-site
USD 140,000 - 210,000
Staff Detection Security Operations Engineer
Staff Detection Security Operations Engineer

Arkenstone Defense • Menlo Park (CA)

On-site
USD 150,000 - 190,000
Competitive Salary
Health and Wellness Programs
Retirement Planning
+3