Observability Platform Engineer

TechDigital Group

Bellevue (WA)

On-site

USD 140,000 - 190,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A leading technology firm in Bellevue, Washington is seeking a candidate to lead the architecture and implementation of an observability strategy within the SIEM modernization ecosystem. Responsibilities include designing telemetry solutions, developing monitoring dashboards, and integrating various monitoring tools. Candidates with experience in managing data events across platforms are encouraged to apply.

Responsibilities

  • Lead the architecture and implementation of a comprehensive observability strategy across the SIEM modernization ecosystem.
  • Design and build end-to-end telemetry and traceability for data events.
  • Develop and maintain dashboards and alerting mechanisms to detect faults and failures.
  • Monitor Blob storage utilization and access patterns to identify issues.
  • Integrate monitoring solutions with Grafana, Azure Monitor, and PowerBI.
  • Ensure observability tooling for Azure Event Hubs: partition health, consumer group lag, throttling events.
  • Integrate monitoring with Grafana, Azure Monitor, and PowerBI for technical, operational, and executive views.
  • Collaborate with Security, Platform, and Data Engineering to drive actionable insights.

Job description

Responsibilities
  • Lead the architecture and implementation of a comprehensive observability strategy across the entire SIEM modernization ecosystem, spanning data pipeline layers (Cribl, Vector, NiFi), event transport (Event Hubs), intermediate storage (Blob), and multiple downstream platforms (Splunk, Snowflake, ADX, Log Analytics, Anvilogic).
  • Design and build end-to-end telemetry and traceability for data events as they move across platforms, enabling real-time visibility into ingestion, transformation, routing, and storage processes.
  • Develop and maintain dashboards and alerting mechanisms to detect:
    • Faults and failures (e.g., dropped messages, ingestion lags, retry loops)
    • Latency or throughput bottlenecks across pipelines
    • Schema mismatches or format errors
    • Duplicate, delayed, or missing data
    • Data quality anomalies at point of ingestion and final storage
  • Instrument each pipeline component (e.g., Cribl workers, Vector agents, NiFi processors) with health and performance metrics, using native exporters, APIs, or custom collectors.
  • Ensure observability tooling is in place for Azure Event Hubs, including partition health, consumer group lag, and throttling events.
  • Monitor Blob storage utilization and access patterns to identify ingest failures, access permission issues, or object lifecycle gaps.
  • Implement and enforce correlation IDs or tracing metadata to follow data across systems and detect where in the pipeline an issue originates.
  • Integrate monitoring solutions with Grafana, Azure Monitor, and PowerBI to support multiple stakeholder needs (technical, operational, and executive-level views).
  • Partner closely with Security Engineering, Platform Engineering, and Data Engineering to ensure observability insights are actionable and result in measurable improvements.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Observability Platform Engineer/DevOps
Observability Platform Engineer/DevOps

Cloud Hybrid Technologies, LLC • Atlanta (GA)

On-site
USD 120,000 - 180,000
Observability Platform Engineer: End-to-End Telemetry
Observability Platform Engineer: End-to-End Telemetry

Cloud Hybrid Technologies, LLC • Atlanta (GA)

On-site
USD 120,000 - 180,000
Senior Observability Engineer
Senior Observability Engineer

Tata Consultancy Services • Los Angeles (CA)

On-site
USD 120,000 - 130,000
Observability Platform Architect
Observability Platform Architect

TechDigital Group • Bellevue (WA)

On-site
USD 140,000 - 190,000
Infrastructure Engineer, Observability
Infrastructure Engineer, Observability

Jobtailor • California (MO)

On-site
USD 140,000 - 190,000
Principal Observability Architect (Splunk & Databricks)
Principal Observability Architect (Splunk & Databricks)

Scicominfra • Atlanta (GA)

On-site
USD 140,000 - 180,000
Health insurance
401(k) retirement plan
Paid time off
Principal Observability Architect (Splunk & Databricks)
Principal Observability Architect (Splunk & Databricks)

scicominfrastructureservices • Atlanta (GA)

On-site
USD 150,000 - 180,000
Health insurance
401(k) matching
Remote work flexibility
Security Data Engineer
Security Data Engineer

TechDigital Group • Overland Park (KS)

On-site
USD 120,000 - 160,000
Observability Architect
Observability Architect

TechDigital Group • Atlanta (GA)

On-site
USD 120,000 - 150,000
Observability Engineer
Observability Engineer

Insight Global • New York (NY)

On-site
USD 120,000 - 180,000