Site Reliability Engineer III — Scale, Observability & Automation

onXmaps, Inc.

Bozeman (MT)

Hybrid

USD 130,000 - 153,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Health benefits including no monthly–$
401(k) matching
Parental leave
Outdoor adventures and Get Out Get Act
Flexible time-away

Job summary

OnXmaps, Inc. is hiring a Site Reliability Engineer to design, build, and maintain the infrastructure enabling our developers to ship reliably at scale.

You will own the infrastructure platform, deployment automation, and observability with IaC, keeping systems performant while enabling fast production access for teams. The role emphasizes Terraform/OpenTofu, Kubernetes, and cloud familiarity, with opportunities to influence architectural decisions and contribute to a distributed,

Qualifications

  • B.S. or M.S. in Computer Science or a related field, or equivalent experience.
  • At least 5+ years of experience with 3+ supporting production systems.
  • Strong interest and experience with Kubernetes, networking, and IaC.
  • Experience with Terraform/OpenTofu.
  • Exposure to at least one major cloud platform.
  • Ability to evaluate technologies based on stability, performance and debuggability.
  • Experience with various data stores (SQL/NoSQL/object storage).

Responsibilities

  • Deploy, monitor and maintain highly available systems using Terraform, CockroachDB and GCP services.
  • Maintain and extend a large Terraform codebase.
  • Analyze systems to improve performance and reduce cost.
  • Automate manual systems to minimize toil.
  • Develop integrations with monitoring/alerting systems (Prometheus, OpenTelemetry, etc.).
  • Drive incident response best practices and participate in on‑call rotations.
  • Collaborate on architectural decisions affecting services and initiatives.

Skills

Kubernetes
Terraform
Cloud platforms
Observability
Infrastructure as Code
Networking
Data stores
Problem solving

Education

B.S./M.S. in Computer Science or related field
5+ years experience with 3+ in production systems

Tools

Terraform
OpenTofu
Kubernetes
GCP
Prometheus

Job description

OnXmaps, Inc. is hiring a Site Reliability Engineer to design, build, and maintain the infrastructure enabling our developers to ship reliably at scale.

You will own the infrastructure platform, deployment automation, and observability with IaC, keeping systems performant while enabling fast production access for teams. The role emphasizes Terraform/OpenTofu, Kubernetes, and cloud familiarity, with opportunities to influence architectural decisions and contribute to a distributed,

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

SRE III — Remote, Scalable Infra & Reliability
SRE III — Remote, Scalable Infra & Reliability

Next Frontier Capital • Colorado

On-site
USD 130,000 - 153,000
No monthly-cost health plan
Equity/stock options
Parental leave
+4
Site Reliability Engineer II: Scalable Infra & Apps
Site Reliability Engineer II: Scalable Infra & Apps

Xometry • Cadentown (KY)

Hybrid
USD 95,000 - 135,000
401(k) match
Medical, dental and vision insurance
Life and disability insurance
+1
Site Reliability Engineer II: Build Scalable Infra
Site Reliability Engineer II: Build Scalable Infra

Xometry • Boston (MA)

Hybrid
USD 135,000 - 165,000
401(k) matching
Medical, dental and vision insurance
Paid time off
+1
Site Reliability Engineer — Platform Observability & Autonomy
Site Reliability Engineer — Platform Observability & Autonomy

MaintainX • San Francisco (CA), Northern (KY)

Hybrid
USD 130,000 - 180,000
Site Reliability Engineer II — Hybrid, AWS & Kubernetes
Site Reliability Engineer II — Hybrid, AWS & Kubernetes

Xometry • Waltham (MA)

Hybrid
USD 135,000 - 155,000
401(k) match
Medical, dental and vision insurance
Life and disability insurance
+4
Site Reliability Engineer II — Scale & Resilience
Site Reliability Engineer II — Scale & Resilience

DAT Freight Solutions • Portland (OR)

Hybrid
USD 95,000 - 134,000
Medical, Dental, Vision
401k matching
Flexible vacation
+1
Senior Observability & SRE Engineer (GCP/Kubernetes)
Senior Observability & SRE Engineer (GCP/Kubernetes)

Ontrac Solutions • New York (NY)

On-site
USD 120,000 - 190,000
Site Reliability Engineer II — Scale, Automate & Observe
Site Reliability Engineer II — Scale, Automate & Observe

DAT • Denver (CO)

Hybrid
USD 95,000 - 134,000
Medical, Dental, Vision
401k matching
Employee Stock Purchase Plan
+3
Site Reliability Engineer III
Site Reliability Engineer III

onXmaps, Inc. • Bozeman (MT)

On-site
USD 130,000 - 153,000
Health benefits including no monthly–$
401(k) matching
Parental leave
+2
Senior Site Reliability Engineer: Scalable Infra & Observability
Senior Site Reliability Engineer: Scalable Infra & Observability

Early Warning • Chicago (IL)

Hybrid
USD 106,000 - 130,000
Healthcare Coverage
401(k) Matching
Paid Time Off
+1