Remote SRE Engineer — Scale, Reliability & Automation

Bright Vision Technologies

Cranberry Township (Butler County)

Remote

USD 100,000 - 180,000

Full time

9 days ago
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Bright Vision Technologies is seeking an experienced Site Reliability Engineer to ensure availability and performance of large-scale distributed systems in production. You will operate at the boundary of development and operations, applying software engineering to infrastructure and automation.

The ideal candidate combines deep systems knowledge with programming skills and a measurement-driven mindset to design, automate, and operate complex services so reliability is engineered, not reactive.

Qualifications

  • Bachelor's degree in Computer Science, Engineering, or a related technical discipline.
  • Ten or more years of SRE, DevOps, or production engineering experience supporting large-scale distributed systems.
  • Strong programming skills in Python, Go, or Java, with the ability to build robust automation and tooling.
  • Hands-on experience operating Linux at scale, including networking and performance tuning.
  • Production experience operating Kubernetes and container-based workloads.
  • Observability tooling knowledge such as Prometheus, Grafana, OpenTelemetry, ELK/EFK, or Datadog.
  • Experience designing and operating CI/CD pipelines for infrastructure and applications.
  • Solid understanding of distributed system design, including consistency models and failure semantics.
  • Demonstrated experience leading incident response and post-incident reviews.
  • Excellent communication and documentation skills.

Responsibilities

  • Define and refine SLOs/SLIs and error budgets for critical services.
  • Lead incident response and post-incident reviews to drive improvements.
  • Design and implement monitoring/logging/tracing strategies (Prometheus, Grafana, OpenTelemetry, ELK/EFK, Datadog).
  • Build on-call processes, runbooks, and escalation paths to reduce toil.
  • Automate operational workflows with Python/Go/Bash.
  • Architect and operate large Kubernetes clusters and container workloads.
  • Design CI/CD pipelines with automated testing, canaries, and progressive rollout.
  • Lead capacity planning and performance engineering with load testing and chaos experiments.
  • Partner with development teams to embed reliability early in design.
  • Strengthen security posture through patch management and secure defaults.
  • Contribute to reliability tooling roadmaps and developer experience.
  • Mentor engineers on SRE practices and blameless culture.

Skills

Python
Go
Java
Linux
Kubernetes
Prometheus
Grafana
OpenTelemetry
CI/CD
Incident response

Education

Bachelor's degree in Computer Science or related

Tools

Datadog
AWS
Azure
GCP
Chaos Monkey
Gremlin
Litmus
Istio

Job description

Bright Vision Technologies is seeking an experienced Site Reliability Engineer to ensure availability and performance of large-scale distributed systems in production. You will operate at the boundary of development and operations, applying software engineering to infrastructure and automation.

The ideal candidate combines deep systems knowledge with programming skills and a measurement-driven mindset to design, automate, and operate complex services so reliability is engineered, not reactive.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Remote Site Reliability Engineer — Scale & Automate
Remote Site Reliability Engineer — Scale & Automate

Bright Vision Technologies • Sterling (VA)

On-site
USD 100,000 - 150,000
Remote SRE — Scale, Resilience & Observability
Remote SRE — Scale, Resilience & Observability

Bright Vision Technologies • United States

On-site
USD 100,000 - 150,000
Competitive base salary
Health benefits
Long-term stability
Senior Site Reliability Engineer - Remote
Senior Site Reliability Engineer - Remote

Bright-Vision-Technologies • United States

Remote
USD 100,000 - 150,000
Remote SRE Technical Lead - Automation & Reliability
Remote SRE Technical Lead - Automation & Reliability

Bright Vision Technologies • Austin (TX), New York (NY)

On-site
USD 100,000 - 150,000
Remote DevOps & SRE Engineer — Cloud Reliability
Remote DevOps & SRE Engineer — Cloud Reliability

Bright Vision Technologies • Columbus (OH), Powell (OH), New Albany (OH), Hilliard (OH)

Remote
USD 100,000 - 150,000
Remote Platform Reliability Engineer – SRE Mastery
Remote Platform Reliability Engineer – SRE Mastery

Bright Vision Technologies • Eden Prairie (MN)

Remote
USD 100,000 - 150,000
Remote DevOps & SRE Engineer for High Reliability
Remote DevOps & SRE Engineer for High Reliability

Socket.dev • Columbus (OH), Powell (OH), New Albany (OH), Hilliard (OH)

On-site
USD 100,000 - 150,000
Remote SRE: Design Resilient, Scalable Systems & Automation
Remote SRE: Design Resilient, Scalable Systems & Automation

OutSolve • Mission (KS)

Remote
USD 90,000 - 130,000
100% remote work environment
Competitive compensation
Professional development opportunities
+1
Senior SRE — Scale, Automation & Uptime
Senior SRE — Scale, Automation & Uptime

Hirebridge • Northern (KY)

Hybrid
USD 110,000 - 145,000
Bonus
Systems Reliability Engineer
Systems Reliability Engineer

Bright Vision Technologies • Sterling (VA)

On-site
USD 100,000 - 150,000