Get more replies from employers
Send a job-specific resume in minutes.
Veriipro is seeking a Senior Observability Engineer to design, implement, and manage enterprise-scale observability platforms. You will lead the rollout of Grafana-based dashboards, from instrumentation to telemetry pipelines, across large environments.
You will migrate from legacy tools, optimize monitoring and logging, and collaborate with Platform Engineering, DevOps, SRE, and Applications teams to raise visibility and reliability.
We are seeking a Senior Observability Engineer with strong experience designing, implementing, and managing enterprise-scale observability platforms. The ideal candidate will have hands-on expertise with Grafana, Terraform, OpenTelemetry, monitoring, logging, alerting, and distributed tracing, along with experience migrating from legacy observability and monitoring platforms.
Design, implement, and maintain enterprise-scale observability and monitoring solutions.
Administer and manage Grafana across large-scale enterprise environments.
Develop and maintain infrastructure using Terraform and Infrastructure as Code (IaC) practices.
Implement and optimize monitoring, alerting, logging, and distributed tracing solutions.
Design and manage OpenTelemetry instrumentation and telemetry pipelines.
Support migrations from platforms such as Splunk, Dynatrace, AppDynamics, New Relic, OpenText OBM, or similar monitoring and observability tools.
Develop dashboards, alerts, metrics, logs, and traces to improve system visibility and operational reliability.
Troubleshoot observability infrastructure and resolve monitoring and telemetry-related issues.
Collaborate with Platform Engineering, DevOps, SRE, Application Development, and Operations teams to establish effective observability standards.
Automate observability platform deployment, configuration, and management using IaC and automation tools.
Establish best practices for scalability, reliability, security, and performance of observability platforms.
Document observability architectures, configurations, standards, and operational procedures.
5+ years of experience in observability, monitoring, operations, SRE, DevOps, or platform engineering.
Hands-on experience administering Grafana in large-scale enterprise environments.
Strong experience with Terraform and Infrastructure as Code (IaC).
Hands-on experience with OpenTelemetry, including instrumentation concepts and telemetry pipelines.
Experience implementing and managing monitoring, alerting, logging, and distributed tracing solutions.
Experience migrating from observability and monitoring platforms such as Splunk, Dynatrace, AppDynamics, New Relic, OpenText OBM, or similar tools.
Strong understanding of enterprise observability architecture and modern telemetry practices.
Experience troubleshooting complex monitoring and observability issues in production environments.
Strong analytical, problem-solving, communication, and collaboration skills.
Experience with cloud platforms such as AWS, Azure, or GCP.
Experience with Kubernetes and containerized environments.
Familiarity with Prometheus, Loki, Tempo, Elasticsearch, or similar observability technologies.
Experience with CI/CD pipelines and automation.
Experience establishing observability standards and best practices across large enterprise environments.