Network Observability/Monitoring Engineer

Insight Global

Dallas, Fort Worth (TX, TX)

On-site

USD 120,000 - 180,000

Full time

7 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Insight Global is seeking a Network Monitoring Platform Engineer to support a large telecommunications client''s customized monitoring and service assurance platform. The role focuses on platform enhancements, upgrades, migrations, troubleshooting, and ongoing operational support in production environments.

The ideal candidate will have Linux administration, experience with Prometheus and Grafana, REST APIs, and automation via Ansible and Jenkins.

Qualifications

  • 5+ years supporting network monitoring or related platforms.
  • Experience building and consuming REST APIs.
  • Proficiency with Prometheus and Grafana.
  • Experience with Ansible, Bash, and Jenkins CI/CD.
  • Production issue troubleshooting and platform maintenance.
  • Experience with relational databases.

Responsibilities

  • Enhance, tune, and maintain an enterprise network monitoring and service assurance platform.
  • Develop and maintain Python-based services and integrations.
  • Troubleshoot production issues and provide operational support as needed.
  • Configure, deploy, upgrade, and manage platform instances across environments.
  • Build and maintain REST API integrations for custom platform functionality.
  • Support monitoring and observability capabilities using Grafana and Prometheus.
  • Develop automation and deployment processes using Ansible, Bash, Jenkins, and CI/CD pipelines.
  • Perform platform upgrades, migrations, and service modernization efforts.
  • Troubleshoot platform performance, data collection, and service health issues.
  • Support Linux-based infrastructure running on Red Hat Enterprise Linux.
  • Work with relational databases to support platform services and integrations.
  • Collaborate with engineering teams to implement new monitoring standards and capabilities.
  • Configure and maintain monitoring services that capture network and device health information from customer environments.

Skills

Networking fundamentals
REST APIs
Prometheus
Grafana
Ansible
Bash scripting
Jenkins

Tools

Prometheus
Grafana
Ansible
Bash scripting
Jenkins

Job description

We are seeking a Network Monitoring Platform Engineer to support and enhance our large telecommunication client's customized network monitoring and service assurance platform. This platform collects device health and performance data for end customers and has been heavily customized to support internal specific policies, protocols, and operational standards.

This role will focus on platform enhancements, upgrades, service migrations, troubleshooting, and ongoing operational support. The ideal candidate will have experience supporting monitoring or observability platforms, strong Linux administration skills, and the ability to develop, enhance, and deploy services in production environments.

Core Responsibilities
  • Enhance, tune, and maintain an enterprise network monitoring and service assurance platform
  • Develop and maintain Python-based services and integrations
  • Troubleshoot production issues and provide operational support as needed
  • Configure, deploy, upgrade, and manage platform instances across environments
  • Build and maintain REST API integrations for custom platform functionality
  • Support monitoring and observability capabilities using Grafana and Prometheus
  • Develop automation and deployment processes using Ansible, Bash, Jenkins, and CI/CD pipelines
  • Perform platform upgrades, migrations, and service modernization efforts
  • Troubleshoot platform performance, data collection, and service health issues
  • Support Linux-based infrastructure running on Red Hat Enterprise Linux
  • Work with relational databases to support platform services and integrations
  • Collaborate with engineering teams to implement new monitoring standards and capabilities
  • Configure and maintain monitoring services that capture network and device health information from customer environments
Required Qualifications
  • 5+ years supporting network monitoring, observability, service assurance, NOC, or telecommunications platforms
  • Experience building and consuming REST APIs
  • Experience with Prometheus and Grafana
  • Experience with Ansible, Bash scripting, and CI/CD pipelines using Jenkins
  • Experience troubleshooting and supporting applications in production environments
  • Experience deploying, configuring, upgrading, and maintaining platform services
  • Understanding of networking fundamentals and how device health and telemetry data are collected and monitored
  • Experience working with relational databases
Preferred Qualifications
  • Experience with Oracle Unified Assurance (Assure1), Netcool, SevOne, ScienceLogic, SolarWinds, LogicMonitor, Splunk, Dynatrace, or similar monitoring platforms
  • Telecommunications, NOC, or OSS/BSS experience
  • Experience with fault management, event correlation, and service assurance concepts
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Network Observability Platform Engineer
Senior Network Observability Platform Engineer

Insight Global • Dallas (TX), Fort Worth (TX)

On-site
USD 120,000 - 180,000
Network Monitoring System Engineer
Network Monitoring System Engineer

Compunnel, Inc. • Baltimore (MD)

On-site
USD 100,000 - 130,000
Network Operations Center (NOC) Engineer
Network Operations Center (NOC) Engineer

Inabia Software & Consulting • Omaha (NE)

On-site
USD 70,000 - 95,000
Network Engineer
Network Engineer

Insight Global • Atlanta (GA)

On-site
USD 90,000 - 130,000
System Administrator (4P 810)
System Administrator (4P 810)

4p-Consulting-Inc. • Atlanta (GA)

On-site
USD 110,000 - 160,000
Senior Enterprise Monitoring and Event Management Engineer
Senior Enterprise Monitoring and Event Management Engineer

Manifest Solutions • Columbus (OH)

On-site
USD 110,000 - 160,000
Network Operations - Tools Engineer
Network Operations - Tools Engineer

Mbi Llc • Newark (NJ)

On-site
USD 80,000 - 100,000
Lead Platform Engineer (SRE)
Lead Platform Engineer (SRE)

InfoVision Inc. • Evansville (IN)

On-site
USD 120,000 - 180,000
Network Engineer - 5072
Network Engineer - 5072

Tier4 Group • United States

On-site
USD 80,000 - 100,000
Senior Observability Engineer – NS2JP00000386
Senior Observability Engineer – NS2JP00000386

Prestige Staffing • Oak Hill (WV)

On-site
USD 100,000 - 130,000