EOG Event Management Triage Engineer

Jobtailor

Austin (TX)

On-site

USD 110,000 - 140,000

Full time

9 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Jobtailor is seeking an experienced IT Operations Specialist to monitor enterprise applications and infrastructure with advanced APM tools, ensuring performance and availability across systems.

You will perform event triage, root-cause analysis, and incident support, collaborating with cloud, network, and application teams, and maintaining operational docs. Requires 5+ years of relevant experience and ability to obtain VA Public Trust clearance. Based in Austin, TX.

Qualifications

  • 5+ years of experience in enterprise IT operations, systems administration, or infrastructure support.
  • 5+ years of experience supporting enterprise applications and platforms, including monitoring and incident management.
  • 5+ years of experience deploying, maintaining, and troubleshooting enterprise-scale applications.
  • 5+ years of monitoring and troubleshooting experience with two or more of: Dynatrace, ScienceLogic, Elastic, AppDynamics, LogicMonitor, SolarWinds.
  • 5+ years of experience supporting Windows, Linux/Unix, and mainframe environments.
  • 1+ year of experience working with AWS, Azure, containers, SaaS, PaaS, or service virtualization technologies.
  • Experience with Microsoft Office Suite (Word, Excel, Teams, PowerPoint).
  • Experience tracking operational performance and supporting incident response activities.
  • Ability to obtain and maintain a VA Public Trust clearance.
  • Must be able to obtain a Personal Identity Verification (PIV) badge.
  • Non-U.S. Nationals must have resided in the United States or U.S. Territories for a minimum of three consecutive years.
  • Must comply with all applicable VA vaccination requirements, if required.
  • Must be willing to complete the VA background investigation process, including the electronic application (eApp), employment history, residential history, and professional references.

Responsibilities

  • Monitor enterprise applications, systems, and infrastructure utilizing advanced monitoring and APM tools.
  • Perform event triage, alert correlation, and root cause analysis of performance and availability issues.
  • Support major incident response efforts by providing real-time monitoring data and technical analysis.
  • Escalate critical alerts and service disruptions to appropriate support teams and stakeholders.
  • Analyze operational trends and recommend improvements to monitoring thresholds and alerting configurations.
  • Collaborate with system administrators, application teams, network engineers, and cloud teams to troubleshoot complex technical issues.
  • Maintain and improve operational documentation and knowledge base articles.
  • Support shift operations, including weekends, holidays, and off-hours as required.

Skills

Root Cause Analysis
Event Triage
Performance Monitoring
Application Performance Monitoring
Enterprise IT Operations
Incident Management
AWS And Azure Experience
Windows Administration
Linux/Unix Administration
Mainframe Support
Service Virtualization

Education

Bachelor's Degree in CS/IT/Engineering
13+ years of relevant professional experience in lieu of a degree

Tools

Dynatrace
ScienceLogic
Elastic
AppDynamics
LogicMonitor
SolarWinds
Microsoft Office Suite

Job description

  • Monitor enterprise applications, systems, and infrastructure utilizing advanced monitoring and APM tools
  • Perform event triage, alert correlation, and root cause analysis of performance and availability issues
  • Support major incident response efforts by providing real-time monitoring data and technical analysis
  • Escalate critical alerts and service disruptions to appropriate support teams and stakeholders
  • Analyze operational trends and recommend improvements to monitoring thresholds and alerting configurations
  • Collaborate with system administrators, application teams, network engineers, and cloud teams to troubleshoot complex technical issues
  • Maintain and improve operational documentation and knowledge base articles
  • Support shift operations, including weekends, holidays, and off-hours as required

Requirements

  • 5+ years of experience in enterprise IT operations, systems administration, or infrastructure support
  • 5+ years of experience supporting enterprise applications and platforms, including monitoring and incident management
  • 5+ years of experience deploying, maintaining, and troubleshooting enterprise-scale applications
  • 5+ years of monitoring and troubleshooting experience with two or more of: Dynatrace, ScienceLogic, Elastic, AppDynamics, LogicMonitor, SolarWinds
  • 5+ years of experience supporting Windows, Linux/Unix, and mainframe environments
  • 1+ year of experience working with AWS, Azure, containers, SaaS, PaaS, or service virtualization technologies
  • Bachelor's Degree in Computer Science, Information Technology, Engineering, or related field and 5+ years of experience; OR 13+ years of relevant professional experience in lieu of a degree
  • Experience with Microsoft Office Suite (Word, Excel, Teams, PowerPoint)
  • Experience tracking operational performance and supporting incident response activities
  • Ability to obtain and maintain a VA Public Trust clearance
  • Must be able to obtain a Personal Identity Verification (PIV) badge
  • Non-U.S. Nationals must have resided in the United States or U.S. Territories for a minimum of three consecutive years
  • Must comply with all applicable VA vaccination requirements, if required
  • Must be willing to complete the VA background investigation process, including the electronic application (eApp), employment history, residential history, and professional references

Core Competencies

Demonstrates expertise in monitoring enterprise applications and systems, performing root cause analysis, and supporting incident management. Proficient in collaborating with cross-functional teams to troubleshoot complex technical issues and improve operational performance.

Highest-signal resume keywords

  • Enterprise IT Operations
  • Monitoring And Troubleshooting
  • Incident Management
  • AWS And Azure Experience
  • Advanced Monitoring Tools

ATS Optimization Keywords

Hard Skills

  • Root Cause Analysis
  • Event Triage
  • Performance Monitoring
  • Application Performance Management
  • Operational Trend Analysis
  • Windows Administration
  • Linux/Unix Administration
  • Mainframe Support
  • Enterprise-Scale Application Deployment
  • Technical Analysis

Soft Skills

  • Collaboration
  • Communication
  • Problem-Solving

Certifications & Qualifications

  • VA Public Trust Clearance
  • Personal Identity Verification (PIV) Badge

Industry Keywords

  • Enterprise Applications
  • Infrastructure Support
  • Operational Documentation
  • Incident Response
  • Service Virtualization

Tools & Technologies

  • Dynatrace
  • ScienceLogic
  • Elastic
  • AppDynamics
  • LogicMonitor
  • SolarWinds
  • Microsoft Office Suite
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Technical Support Advisor
Technical Support Advisor

Jobtailor • Franklin (TN)

On-site
USD 90,000 - 120,000
Enterprise Monitoring Engineer
Enterprise Monitoring Engineer

Jobtailor • Billings (MT)

On-site
USD 90,000 - 130,000
Senior Engineer, Applications Systems
Senior Engineer, Applications Systems

Jobtailor • Town of Florida (NY)

On-site
USD 120,000 - 160,000
IT Platform Admin I – IV
IT Platform Admin I – IV

Jobtailor • City of Utica (NY)

On-site
USD 90,000 - 120,000
Senior Systems Administrator – Hybrid
Senior Systems Administrator – Hybrid

Jobtailor • Virginia (MN)

On-site
USD 120,000 - 210,000
Senior Infrastructure Services Analyst
Senior Infrastructure Services Analyst

Jobtailor • Atlanta (GA)

On-site
USD 120,000 - 180,000
Senior System Administrator – Hybrid
Senior System Administrator – Hybrid

Jobtailor • Bethesda (MD)

On-site
USD 120,000 - 170,000
Infrastructure Systems Engineer II
Infrastructure Systems Engineer II

Jobtailor • Town of New Castle (NY)

On-site
USD 110,000 - 160,000
Senior System Reliability & Support Specialist – Production Support
Senior System Reliability & Support Specialist – Production Support

Jobtailor • Alabama

On-site
USD 60,000 - 80,000
Senior Manager, Technology Operations Support
Senior Manager, Technology Operations Support

Jobtailor • Columbus (OH)

Hybrid
USD 110,000 - 170,000