Site Reliability Engineer III — Scale, Observability & Automation

onXmaps, Inc.

Bozeman (MT)

Hybrid

USD 130,000 - 153,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health benefits including no monthly–$
401(k) matching
Parental leave
Outdoor adventures and Get Out Get Act
Flexible time-away

Job summary

OnXmaps, Inc. is hiring a Site Reliability Engineer to design, build, and maintain the infrastructure enabling our developers to ship reliably at scale.

You will own the infrastructure platform, deployment automation, and observability with IaC, keeping systems performant while enabling fast production access for teams. The role emphasizes Terraform/OpenTofu, Kubernetes, and cloud familiarity, with opportunities to influence architectural decisions and contribute to a distributed,

Qualifications

  • B.S. or M.S. in Computer Science or a related field, or equivalent experience.
  • At least 5+ years of experience with 3+ supporting production systems.
  • Strong interest and experience with Kubernetes, networking, and IaC.
  • Experience with Terraform/OpenTofu.
  • Exposure to at least one major cloud platform.
  • Ability to evaluate technologies based on stability, performance and debuggability.
  • Experience with various data stores (SQL/NoSQL/object storage).

Responsibilities

  • Deploy, monitor and maintain highly available systems using Terraform, CockroachDB and GCP services.
  • Maintain and extend a large Terraform codebase.
  • Analyze systems to improve performance and reduce cost.
  • Automate manual systems to minimize toil.
  • Develop integrations with monitoring/alerting systems (Prometheus, OpenTelemetry, etc.).
  • Drive incident response best practices and participate in on‑call rotations.
  • Collaborate on architectural decisions affecting services and initiatives.

Skills

Kubernetes
Terraform
Cloud platforms
Observability
Infrastructure as Code
Networking
Data stores
Problem solving

Education

B.S./M.S. in Computer Science or related field
5+ years experience with 3+ in production systems

Tools

Terraform
OpenTofu
Kubernetes
GCP
Prometheus

Job description

OnXmaps, Inc. is hiring a Site Reliability Engineer to design, build, and maintain the infrastructure enabling our developers to ship reliably at scale.

You will own the infrastructure platform, deployment automation, and observability with IaC, keeping systems performant while enabling fast production access for teams. The role emphasizes Terraform/OpenTofu, Kubernetes, and cloud familiarity, with opportunities to influence architectural decisions and contribute to a distributed,

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer - Kubernetes & Terraform
Senior Site Reliability Engineer - Kubernetes & Terraform

onX • Bozeman (MT)

Hybrid
USD 130,000 - 153,000
Health benefits
Parental leave
401k matching
+3
Senior Cloud Observability & SRE Engineer
Senior Cloud Observability & SRE Engineer

Ontrac Solutions • Chicago (IL)

On-site
USD 120,000 - 180,000
Observability & SRE Engineer — Cloud, Kubernetes, & Automation
Observability & SRE Engineer — Cloud, Kubernetes, & Automation

Ontrac Solutions • Arizona

Hybrid
USD 140,000 - 190,000
Verification cost reimbursement
Site Reliability Engineer III
Site Reliability Engineer III

onXmaps, Inc. • Bozeman (MT)

Hybrid
USD 130,000 - 153,000
Health benefits including no monthly–$
401(k) matching
Parental leave
+2
Senior Site Reliability Engineer: Cloud Automation & Scale
Senior Site Reliability Engineer: Cloud Automation & Scale

Okta • New York (NY)

On-site
USD 174,000 - 239,000
Senior Site Reliability Engineer: Scalable Infra & Observability
Senior Site Reliability Engineer: Scalable Infra & Observability

Early Warning • Chicago (IL)

Hybrid
USD 106,000 - 130,000
Healthcare Coverage
401(k) Matching
Paid Time Off
+1
Senior Site Reliability Engineer — Observability & Automation
Senior Site Reliability Engineer — Observability & Automation

PVH (Tommy Hilfiger/Calvin Klein) • Salt Lake City (UT)

Hybrid
USD 130,000 - 160,000
Equity
Healthcare
Retirement plan
+3
Site Reliability Engineer — Build Ultra-Reliable Systems
Site Reliability Engineer — Build Ultra-Reliable Systems

Virtual Tech Gurus • Puerto Rico

On-site
USD 140,000 - 210,000
Site Reliability Engineer
Site Reliability Engineer

NextGen | GTA: A Kelly Telecom Company • Mount Laurel Township (NJ)

On-site
USD 110,000 - 170,000
Senior Observability & SRE Engineer — GCP/Kubernetes
Senior Observability & SRE Engineer — GCP/Kubernetes

Ontrac Solutions • United States

On-site
USD 120,000 - 180,000