Orlando, United States | Posted on 12/01/2025
The Sr. ObservabilityEngineer is responsible for designing, deploying, and optimizing client'senterprise observability ecosystem. This role delivers hands-on implementationand consulting expertise, focusing on LogicMonitor and modern observability practicesto drive actionable insights, predictive analytics, and operational excellenceacross infrastructure, network, and application layers.
Key Responsibilities
- Deploy, configure, and optimize LogicMonitor forenterprise-scale observability.
- Design and build custom dashboards for actionableinsights and performance monitoring.
- Implement and manage data analytics workflows,including advanced scripting in Python for automation and reporting.
- Integrate and manage data pipelines leveraging Kafkaand related streaming technologies.
- Ensure seamless data flow into Grafana forvisualization and monitoring.
- Develop and maintain integrations betweenobservability, ITSM (ServiceNow), and event management tools (PagerDuty,Slack, BigPanda).
- Standardize alert thresholds, escalation paths, andtelemetry mappings across global regions.
- Define and maintain event, alert, and rule logic toensure accurate correlation and minimal noise.
- Manage data ingestion pipelines from SNMP, syslog,APIs, and third-party sources into LogicMonitor and downstream analyticssystems.
- Advise on and implement AI-driven observability tools,including Amazon Bedrock, to enhance predictive analytics and anomalydetection.
- Partner with network, server, and application teams tovalidate data flows, performance metrics, and dependency mapping.
- Automate configuration and onboarding processes via APIand scripting (Python, PowerShell, REST).
- Support incident and problem management teams bycorrelating events across multiple tools to accelerate root causeanalysis.
- Document integrations, processes, and governance modelsfor sustained operational excellence.
- Serve as technical SME supporting observability toolupgrades, testing, and cross-platform enhancements.
- Collaborate with stakeholders to align monitoringstrategies with business objectives.
Core Expertise &Skills
- LogicMonitor platform deployment, configuration, andoptimization.
- Deep understanding of observability frameworks, bestpractices, and enterprise monitoring.
- Python scripting for data analytics, automation, andadvanced reporting.
- Kafka-based data streaming and integration.
- Grafana dashboarding and visualization.
- Experience with AI platforms and emerging observabilitytechnologies, including Amazon Bedrock.
- Familiarity with ITSM and event management systems(ServiceNow, PagerDuty, BigPanda).
- Telemetry protocols (SNMP, syslog, NetFlow, APIs) anddata flow architecture.
- Strong understanding of alerting, correlation logic,and performance baselining.
- Excellent analytical, documentation, and communicationskills; ability to work across engineering and operations teams.
- 5-10years of experience in network monitoring, observability, orinfrastructure engineering roles.
- Hands-on experience with LogicMonitor, Splunk,ThousandEyes, Datadog, Dynatrace, Cisco DNAC, and related platforms.
- Scripting and automation experience (Python,PowerShell, REST APIs)
- Looking for an expert in the Logic Monitor platform
- Customer team is decommissioning 4 platforms intoLogicMonitor