Remote Observability Engineer — Scale Metrics & Tracing

Bright Vision Technologies

South Windsor (CT)

On-site

USD 100,000 - 160,000

Full time

5 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Bright Vision Technologies is seeking an Observability Engineer to design and operate enterprise-grade observability platforms for metrics, logs, and traces. This remote, U.S.-based role focuses on scalable deployments, signal quality, and on-call readiness across engineering teams.

You will drive OpenTelemetry adoption, define SLOs/SLIs, and collaborate with SRE and platform teams to integrate observability into deployment pipelines.

Qualifications

  • Bachelor’s degree in Computer Science or a related field.
  • 12+ years of experience in SRE, platform engineering, or observability roles.
  • Hands-on experience with Prometheus, Grafana, and at least one commercial observability platform such as Datadog, New Relic, or Splunk.
  • Strong understanding of OpenTelemetry, distributed tracing, and structured logging.
  • Proficiency in Go, Python, or Java.
  • Experience operating high-cardinality, high-throughput metrics and log pipelines.
  • Strong understanding of SLOs, error budgets, and SRE principles.
  • Experience integrating observability with CI/CD and incident management tooling.
  • Solid grasp of Linux internals, networking, and container platforms.
  • Excellent communication and collaboration skills.

Responsibilities

  • Design and operate enterprise-grade observability platforms covering metrics, logs, traces, events, and synthetic monitoring.
  • Architect Prometheus / Thanos / Mimir, Grafana, Loki, Tempo, OpenTelemetry, and Datadog deployments for high availability and scale.
  • Develop standards for service instrumentation, including OpenTelemetry adoption, metric naming, label cardinality, and structured logging conventions.
  • Define and enforce SLOs, SLIs, and error budgets, and build the dashboards and alerts that operationalize them.
  • Build alerting strategies that minimize noise, surface actionable signals, and integrate cleanly with on-call workflows in PagerDuty, Opsgenie, or similar tools.
  • Operate large-scale time-series and log storage platforms, balancing retention, query performance, and cost.
  • Design distributed tracing pipelines and help teams use traces to diagnose latency and reliability issues.
  • Develop self-service tooling, paved-road libraries, and templates that make adoption of observability standards easy for product teams.
  • Drive cost management and label-cardinality discipline across the observability estate.
  • Lead incident response readiness improvements through better dashboards, alerting hygiene, and post-incident analysis tooling.
  • Partner with SRE and platform teams to integrate observability into deployment pipelines, canary analysis, and progressive delivery workflows.
  • Evaluate and recommend observability vendors and open-source tools based on cost, capability, and operational maturity.
  • Mentor engineering teams on observability fundamentals, debugging techniques, and SLO-driven operations.
  • Maintain documentation, onboarding guides, and runbooks for the observability platform.

Skills

Prometheus
Grafana
OpenTelemetry
Distributed tracing
Structured logging
Go
Python
Java
CI/CD
SRE principles

Education

Bachelor's degree in Computer Science or related field

Tools

Datadog
New Relic
Splunk
Thanos

Job description

Bright Vision Technologies is seeking an Observability Engineer to design and operate enterprise-grade observability platforms for metrics, logs, and traces. This remote, U.S.-based role focuses on scalable deployments, signal quality, and on-call readiness across engineering teams.

You will drive OpenTelemetry adoption, define SLOs/SLIs, and collaborate with SRE and platform teams to integrate observability into deployment pipelines.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote Observability Engineer - Scale Metrics & Tracing
Remote Observability Engineer - Scale Metrics & Tracing

Bright-Vision-Technologies • United States

Remote
USD 89,000 - 110,000
Remote Observability Engineer: Scale, Metrics & Tracing
Remote Observability Engineer: Scale, Metrics & Tracing

Bright Vision Technologies • Farmington Hills (MI)

On-site
USD 89,000 - 110,000
Senior Observability Engineer (Remote)
Senior Observability Engineer (Remote)

Bright-Vision-Technologies • United States

Remote
USD 75,000 - 85,000
Remote Monitoring Engineer — Observability & SRE
Remote Monitoring Engineer — Observability & SRE

Bright Vision Technologies • Palo Alto (CA)

On-site
USD 75,000 - 85,000
Remote Telemetry Engineer - Observability/SRE Expert
Remote Telemetry Engineer - Observability/SRE Expert

Bright Vision Technologies • Charlotte (NC)

On-site
USD 100,000 - 150,000
Remote Telemetry Engineer: Observability & SRE
Remote Telemetry Engineer: Observability & SRE

Bright-Vision-Technologies • United States

Remote
USD 100,000 - 150,000
Remote Systems Observability Specialist
Remote Systems Observability Specialist

Bright-Vision-Technologies • United States

Remote
USD 135,000 - 155,000
Remote Observability Platform Engineer — Scale Telemetry
Remote Observability Platform Engineer — Scale Telemetry

BairesDev • Peru (IL)

On-site
USD 120,000 - 170,000
100% remote work
USD or local currency pay
Home office setup
+3
Senior Observability Platform Engineer - Scale & Reliability
Senior Observability Platform Engineer - Scale & Reliability

CVS Health • West Virginia

Hybrid
USD 83,000 - 222,000
Observability Leader: Enterprise Telemetry & SRE
Observability Leader: Enterprise Telemetry & SRE

Truist • Charlotte (NC)

On-site
USD 140,000 - 190,000
Medical insurance
Dental insurance
Vision insurance
+2