AI Ops Engineer

Hewlett Packard Enterprise Company in

San Juan (PR)

Hybrid

USD 110,000 - 160,000

Full time

2 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Relocation support
Health & wellbeing benefits
Professional development programs

Job summary

Hewlett Packard Enterprise is seeking an AI Ops Engineer to design intelligent monitoring and automate incident response for cloud and on‑prem environments. You will build data pipelines, develop predictive models, and collaborate with DevOps, SRE, and security teams to improve platform reliability.

Ideal candidates hold a CS/IT degree and 5–8 years of experience in IT operations or AI‑driven automation, with strong Python and observability tool knowledge (Splunk, Datadog, Prometheus, Grafana).

Qualifications

  • Bachelor's degree in Computer Science, Information Technology, or a related field.
  • 5–8 years of experience in IT operations, reliability engineering, platform engineering, or AI-driven operations and automation.
  • Proficiency in Python for automation, data processing, and tooling.
  • Experience with observability platforms (Splunk, Datadog, Prometheus, Grafana).
  • Knowledge of cloud platform operations (Azure, AWS, GCP) is preferred.
  • Knowledge of ML model lifecycle management and deployment is advantageous.
  • Experience with IT service management and workflow orchestration tools such as ServiceNow, PagerDuty, or Rundeck is beneficial.

Responsibilities

  • Design and maintain intelligent monitoring and incident response workflows across cloud and on‑premises systems.
  • Build and optimize data pipelines to collect, normalize, and enrich logs, metrics, and events for operational analytics.
  • Develop and deploy predictive models to detect anomalies, forecast outages, and reduce mean time to resolution.
  • Automate remediation runbooks and integrate alerting with service management platforms to enhance system reliability.
  • Collaborate with DevOps, Site Reliability, and security teams to improve platform performance, availability, and governance.

Skills

Python automation
Reliability engineering
AI-driven operations
Observability mindset

Education

Bachelor's degree in Computer Science/IT or related field

Tools

Splunk
Datadog
Prometheus
Grafana
ServiceNow
PagerDuty
Rundeck

Job description

This role has been designed as 'Hybrid' with a requirement that you will work on average 2 days per week from an HPE office.

Who We Are:

Hewlett Packard Enterprise is the global edge-to-cloud company advancing the way people live and work. We help companies connect, protect, analyze, and act on their data and applications wherever they live, from edge to cloud, so they can turn insights into outcomes at the speed required to thrive in today's complex world. Our culture thrives on finding new and better ways to accelerate what's next. We know varied backgrounds are valued and succeed here. We have the flexibility to manage our work and personal needs. We make bold moves, together, and are a force for good. If you are looking to stretch and grow your career our culture will embrace you. Open up opportunities with HPE.

Job Description:
Job Family Definition:

Designs, develops, troubleshoots, and debugs software programs for enhancements and new product development. Develops software components such as operating systems, compilers, routers, networks, utilities, databases, and Internet-related tools. Assesses hardware compatibility and/or influences hardware design decisions.

Applies specialized subject matter expertise to resolve common and occasionally complex technical challenges, recommending alternatives as needed. May act as a project lead and provide guidance to junior professionals. Exercises independent judgment and collaborates with others to determine the most effective methods for accomplishing work and achieving objectives.

Responsibilities:
  • Design and maintain intelligent monitoring and incident response workflows across both cloud and on-premises systems.
  • Build and optimize data pipelines to collect, normalize, and enrich logs, metrics, and events for operational analytics.
  • Develop and deploy predictive models to detect anomalies, forecast outages, and reduce mean time to resolution.
  • Automate remediation runbooks and integrate alerting with service management platforms to enhance system reliability.
  • Collaborate with DevOps, Site Reliability, and security teams to improve platform performance, availability, and operational governance.
AI Ops Engineer (Finance)

AI Ops Engineer

This role has been designed as 'Hybrid' with a requirement that you will work on average 2 days per week from an HPE office.

Who We Are:

Hewlett Packard Enterprise is the global edge-to-cloud company advancing the way people live and work. We help companies connect, protect, analyze, and act on their data and applications wherever they live, from edge to cloud, so they can turn insights into outcomes at the speed required to thrive in today's complex world. Our culture thrives on finding new and better ways to accelerate what's next. We know varied backgrounds are valued and succeed here. We have the flexibility to manage our work and personal needs. We make bold moves, together, and are a force for good. If you are looking to stretch and grow your career our culture will embrace you. Open up opportunities with HPE.

Job Description:
Job Family Definition:

Designs, develops, troubleshoots, and debugs software programs for enhancements and new product development. Develops software components such as operating systems, compilers, routers, networks, utilities, databases, and Internet-related tools. Assesses hardware compatibility and/or influences hardware design decisions.

Management Level Definition:

Applies specialized subject matter expertise to resolve common and occasionally complex technical challenges, recommending alternatives as needed. May act as a project lead and provide guidance to junior professionals. Exercises independent judgment and collaborates with others to determine the most effective methods for accomplishing work and achieving objectives.

Responsibilities:
  • Design and maintain intelligent monitoring and incident response workflows across both cloud and on-premises systems.
  • Build and optimize data pipelines to collect, normalize, and enrich logs, metrics, and events for operational analytics.
  • Develop and deploy predictive models to detect anomalies, forecast outages, and reduce mean time to resolution.
  • Automate remediation runbooks and integrate alerting with service management platforms to enhance system reliability.
  • Collaborate with DevOps, Site Reliability, and security teams to improve platform performance, availability, and operational governance.
Education and Experience Required:
  • Bachelor's degree in Computer Science, Information Technology, or a related field.
  • 5-8 years of experience in IT operations, reliability engineering, platform engineering, or roles focused on AI-driven operations and automation.
Knowledge and Skills:
  • Proficiency in Python for automation, data processing, and operational tooling - required.
  • Experience with observability platforms such as Splunk, Datadog, Prometheus, or Grafana - required.
  • Familiarity with cloud platform operations (Azure, AWS, or Google Cloud) - preferred.
  • Knowledge of machine learning model lifecycle management and deployment - advantageous.
  • Experience with IT service management and workflow orchestration tools such as ServiceNow, PagerDuty, or Rundeck - beneficial.

Relocation support is provided for eligible candidates

What We Can Offer You:
Health & Wellbeing

We strive to provide our team members and their loved ones with a comprehensive suite of benefits that supports their physical, financial and emotional wellbeing.

Personal & Professional Development

We also invest in your career because the better you are, the better we all are. We have specific programs catered to helping you reach any career goals you have - whether you want to become a knowledge expert in your field or apply your skills to another division.

Unconditional Inclusion

We are unconditionally inclusive in the way we work and celebrate individual uniqueness. We know varied backgrounds are valued and succeed here. We have the flexibility to manage our work and personal needs. We make bold moves, together, and are a force for good.

Let's Stay Connected:

Follow @HPECareers on Instagram to see the latest on people, culture and tech at HPE.

#puertorico

#networking

HPE is an Equal Employment Opportunity/ Veterans/Disabled/LGBT employer. We do not discriminate on the basis of race, gender, or any other protected category, and all decisions we make are made on the basis of qualifications, merit, and business need. Our goal is to be one global team that is representative of our customers, in an inclusive environment where we can continue to innovate and grow together.

Hewlett Packard Enterprise is EEO Protected Veteran/ Individual with Disabilities.

HPE will comply with all applicable laws related to employer use of arrest and conviction records, including laws requiring employers to consider for employment qualified applicants with criminal histories.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Ops Engineer
AI Ops Engineer

Hewlett Packard Enterprise Development LP • San Juan (PR)

Hybrid
USD 110,000 - 170,000
Principal Engineer – Generative AI & LLM Platforms
Principal Engineer – Generative AI & LLM Platforms

Hewlett Packard Enterprise Development LP • San Juan (PR)

Hybrid
USD 120,000 - 190,000
Engineering Tools Support Engineer
Engineering Tools Support Engineer

Hewlett Packard Enterprise • San Juan (PR)

Hybrid
USD 70,000 - 100,000
Health & Wellbeing
Personal & Professional Development
Unconditional Inclusion
User Interface / Front End Engineer
User Interface / Front End Engineer

Hewlett Packard Enterprise Development LP • San Juan (PR)

Hybrid
USD 100,000 - 150,000
Relocation support
User Interface / Front End Engineer
User Interface / Front End Engineer

Socket.dev • San Juan (PR)

Hybrid
USD 80,000 - 120,000
Relocation support
Health benefits
Career development
Engineering Tools Support Engineer
Engineering Tools Support Engineer

Hewlett Packard Enterprise Company • Friday Harbor (WA)

Hybrid
USD 90,000 - 120,000
Health & Wellbeing
Career development
Unconditional Inclusion
AI & ML Software Engineer
AI & ML Software Engineer

Hewlett Packard Enterprise Development LP • Fort Collins (CO)

Hybrid
USD 144,000 - 273,000
Health & wellbeing
Career development
Inclusive culture
Principal Site Reliability Engineer
Principal Site Reliability Engineer

Hewlett Packard Enterprise • San Juan (PR)

Hybrid
USD 120,000 - 160,000
Comprehensive benefits suite
Personal & professional development opportunities
Unconditional inclusion in the workplace
Engineering Tools Support Engineer
Engineering Tools Support Engineer

Hewlett Packard Enterprise Company in • San Juan (PR)

Hybrid
USD 70,000 - 100,000
Software Engineering Manager – Generative AI & Enterprise Platforms
Software Engineering Manager – Generative AI & Enterprise Platforms

Hewlett Packard Enterprise Development LP • San Juan (PR)

Hybrid
USD 190,000 - 260,000
Health & Wellbeing
Career Development Programs
Unconditional Inclusion