Remote Site Observability Engineer: Scale & Insight

Bright Vision Technologies

Issaquah (WA)

On-site

USD 100,000 - 150,000

Full time

7 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Bright Vision Technologies is seeking a Site Observability Engineer to design and operate scalable observability platforms spanning metrics, logs, traces, events, and dashboards. You will translate telemetry into actionable insights for engineers and business stakeholders.

The role includes shaping SLOs/SLIs, implementing OpenTelemetry standards, and deploying Prometheus, Grafana, Loki, Tempo, and Datadog at scale.

Qualifications

  • Bachelor’s degree in Computer Science or a related field.
  • Five or more years of experience in SRE, platform engineering, or observability roles.
  • Deep hands-on experience with Prometheus, Grafana, and at least one major commercial observability platform such as Datadog, New Relic, or Splunk.
  • Strong understanding of OpenTelemetry, distributed tracing, and structured logging.
  • Proficiency in Go, Python, or Java.
  • Experience operating high-cardinality, high-throughput metrics and log pipelines.
  • Strong understanding of SLOs, error budgets, and SRE principles.
  • Experience integrating observability with CI/CD and incident management tooling.
  • Solid grasp of Linux internals, networking, and container platforms.
  • Excellent communication and collaboration skills.

Responsibilities

  • Design and operate enterprise-grade observability platforms covering metrics, logs, traces, events, and synthetic monitoring.
  • Architect Prometheus / Thanos / Mimir, Grafana, Loki, Tempo, OpenTelemetry, and Datadog deployments for high availability and scale.
  • Develop standards for service instrumentation, including OpenTelemetry adoption, metric naming, label cardinality, and structured logging conventions.
  • Define and enforce SLOs, SLIs, and error budgets, and build dashboards and alerts that operationalize them.
  • Build alerting strategies that minimize noise and integrate with on-call workflows.
  • Operate large-scale time-series and log storage platforms balancing retention, query performance, and cost.
  • Design distributed tracing pipelines and help teams diagnose latency and reliability issues.
  • Develop self-service tooling and templates to simplify adoption of observability standards.
  • Drive cost management and label-cardinality discipline across the observability estate.
  • Lead incident response readiness improvements through dashboards and post-incident analysis tooling.
  • Partner with SRE and platform teams to integrate observability into deployment pipelines and progressive delivery workflows.
  • Evaluate observability vendors and open-source tools based on cost and capability.
  • Mentor engineering teams on observability fundamentals and SLO-driven operations.
  • Maintain documentation, onboarding guides, and runbooks for the observability platform.

Skills

SRE principles
Observability
Prometheus
Grafana
OpenTelemetry
Distributed tracing
Go
Linux internals
CI/CD integration
Communication

Education

Bachelor's degree in Computer Science or related field

Tools

Datadog
New Relic
Splunk
Thanos
Mimir
Cortex
Loki
Tempo

Job description

Bright Vision Technologies is seeking a Site Observability Engineer to design and operate scalable observability platforms spanning metrics, logs, traces, events, and dashboards. You will translate telemetry into actionable insights for engineers and business stakeholders.

The role includes shaping SLOs/SLIs, implementing OpenTelemetry standards, and deploying Prometheus, Grafana, Loki, Tempo, and Datadog at scale.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote Senior Observability Engineer – Scale & Insights
Remote Senior Observability Engineer – Scale & Insights

Bright Vision Technologies • Nashua (NH)

On-site
USD 100,000 - 160,000
Remote Observability Engineer: Scale, Metrics & Tracing
Remote Observability Engineer: Scale, Metrics & Tracing

Bright Vision Technologies • Farmington Hills (MI)

On-site
USD 89,000 - 110,000
Remote Monitoring Engineer — Observability & SRE
Remote Monitoring Engineer — Observability & SRE

Bright Vision Technologies • Palo Alto (CA)

On-site
USD 75,000 - 85,000
Senior Observability Engineer - Remote
Senior Observability Engineer - Remote

Bright Vision Technologies • Round Rock (TX)

On-site
USD 135,000 - 155,000
Remote Telemetry Engineer - Observability/SRE Expert
Remote Telemetry Engineer - Observability/SRE Expert

Bright Vision Technologies • Charlotte (NC)

On-site
USD 100,000 - 150,000
Senior Observability & Reliability Engineer - Remote
Senior Observability & Reliability Engineer - Remote

Bright Vision Technologies • Sterling (VA)

On-site
USD 100,000 - 150,000
Remote Observability Platform Engineer — Scale Telemetry
Remote Observability Platform Engineer — Scale Telemetry

BairesDev • Peru (IL)

On-site
USD 120,000 - 170,000
100% remote work
USD or local currency pay
Home office setup
+3
Senior Cloud Observability & SRE Engineer
Senior Cloud Observability & SRE Engineer

Ontrac Solutions • Chicago (IL)

On-site
USD 120,000 - 180,000
Observability Engineer: Scale Telemetry & AI Reliability
Observability Engineer: Scale Telemetry & AI Reliability

Baseten • New York (NY)

On-site
USD 165,000 - 330,000
Competitive compensation including 1x/
Full medical/dental/vision insurance
Flexible PTO with Winter Break
+3
Site Reliability & Observability Engineer
Site Reliability & Observability Engineer

Stability Technology • United States

On-site
USD 120,000 - 170,000