AI and Machine Learning Engineer

Hewlett Packard Enterprise Company

Springs (NY)

Hybrid

USD 121,000 - 277,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health & Wellbeing
Professional development
Unconditional inclusion

Job summary

Hewlett Packard Enterprise Company is seeking an AI and Machine Learning Engineer to advance high-performance computing and AI workloads within a hybrid work model, involving two days per week in an HPE office. The role focuses on deploying scalable AI systems, optimizing transformers, and coordinating with software/hardware partners to enhance performance across DL workloads.

You will contribute to AI/ML benchmarks, containerized environments, and data-centric computing, while communicating

Qualifications

  • Master's degree or PhD in Computer Science, Engineering, Information Technology or Systems, or relevant field.
  • 5+ years of experience in ML/AI and HPC environments.
  • Experience with NCCL, HPL benchmarks and AI workloads.

Responsibilities

  • Install and configure complex IT infrastructure components (servers, storage, network).
  • Develop scripts and configurations for automating deployment.
  • Study and improve performance of Large Language Models on HPE GPU servers.
  • Analyze server workloads and optimize DL/ML code and hardware use.
  • Document guidance and reports on AI workload and model selection.

Skills

ML/AI experience
HPC
Python/C++
Transformer models
CI/CD

Education

Master's or PhD in CS/Engineering/IT or related field

Tools

Slurm
NCCL
Lustre
Weka I/O
Docker/Kubernetes

Job description

AI and Machine Learning Engineer

This role has been designed as 'Hybrid' with a requirement that you will work on average 2 days per week from an HPE office.

Who We Are:

Hewlett Packard Enterprise is the global edge-to-cloud company advancing the way people live and work. We help companies connect, protect, analyze, and act on their data and applications wherever they live, from edge to cloud, so they can turn insights into outcomes at the speed required to thrive in today's complex world. Our culture thrives on finding new and better ways to accelerate what's next. We know varied backgrounds are valued and succeed here. We have the flexibility to manage our work and personal needs. We make bold moves, together, and are a force for good. If you are looking to stretch and grow your career our culture will embrace you. Open up opportunities with HPE.

Job Description:

High Performance Computing, AI and Labs are a critical element of HPE. We are focused on delivering innovative solutions that accelerate our customers' digital transformation, enabling them to tackle their complex, and data-intensive workloads. Combining deep expertise and the development of the world's most cutting-edge, high-performance supercomputers, is defining the next era of computing delivering valuable insight & innovation. Join us and redefine what's next for you.

Responsibilities:
  • Installs and configures complex IT infrastructure components (servers, storage, network)
  • Develop software scripts and configurations for automating deployment.
  • Study and improve the performance of Large Language Models run on HPE GPU servers
  • Performs system level analysis of server workloads on various HPE platforms running DL and ML code to include accelerated hardware and high-speed networks like InfiniBand
  • Writes white papers and other guidance documents for AI workload and model selection
  • Captures and reviews system performance data, logs, traces to understand workload behaviour
  • Develops software and scripts that help analyse AI workload performance data
  • Communicates technical work well and can provide summaries of work to non-technical colleagues
  • Works with software and hardware partners in optimizing systems and resolving performance issues
  • Documents and reports issues discovered when testing and evaluating the systems
  • Communicates project status and concerns to management in a timely manner
  • Provides guidance to less-experienced staff members.
  • Runs AI and HPC benchmarks.
Education and Experience Required:
  • Master's degree or PhD in Computer Science, Engineering, Information Technology or Systems, or relevant field.
  • 5+ years of experience.
Knowledge and Skills:
  • 5+ years of experience in Machine Learning/Artificial Intelligence and 5+ years of experience in HPC
  • Experience running NCCL, HPL and AI benchmarks.
  • Experience working with containers and distributed deep learning and neural networks, to include transformers used in generative AI projects
  • Experience working with High Performance Computer Servers, High Performance Networking, and associated software, including resource managers like Slurm
  • Experience working with Weka I/O, NFTS and Lustre File Systems
  • Programming experience in Python, C, C++
  • Strong analytical and critical thinking skills
  • Scripting, process automation and CI/CD are strongly desired
  • Must be a self-starter and be able to work with minimum supervision in a semi-remote setting

Artificial Intelligence Technologies, Cross Domain Knowledge, Data Engineering, Data Science, Design Thinking, Development Fundamentals, Full Stack Development, IT Performance, Machine Learning Operations, Scalability Testing, Security-First Mindset.

What We Can Offer You:
Health & Wellbeing

We strive to provide our team members and their loved ones with a comprehensive suite of benefits that supports their physical, financial and emotional wellbeing.

Personal & Professional Development

We also invest in your career because the better you are, the better we all are. We have specific programs catered to helping you reach any career goals you have - whether you want to become a knowledge expert in your field or apply your skills to another division.

Unconditional Inclusion

We are unconditionally inclusive in the way we work and celebrate individual uniqueness. We know varied backgrounds are valued and succeed here. We have the flexibility to manage our work and personal needs. We make bold moves, together, and are a force for good.

Let's Stay Connected:

Follow @HPECareers on Instagram to see the latest on people, culture and tech at HPE.

#unitedstates

Job:

Engineering

Job Level:

TCP_03

The expected salary/wage range for this position is provided below. Actual offer may vary from this range based upon geographic location, work experience, education/training, and/or skill level.
- United States of America: Annual Salary USD 120,500 - 276,500 in Texas
The listed salary range reflects base salary. Variable incentives may also be offered.

Information about employee benefits offered in the US can be found at https://myhperewards.com/main/new-hire-enrollment.html

HPE is an Equal Employment Opportunity/ Veterans/Disabled/LGBT employer. We do not discriminate on the basis of race, gender, or any other protected category, and all decisions we make are made on the basis of qualifications, merit, and business need. Our goal is to be one global team that is representative of our customers, in an inclusive environment where we can continue to innovate and grow together. Please click here: Equal Employment Opportunity .

Hewlett Packard Enterprise is EEO Protected Veteran/ Individual with Disabilities.

HPE will comply with all applicable laws related to employer use of arrest and conviction records, including laws requiring employers to consider for employment qualified applicants with criminal histories.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI and Machine Learning Engineer
AI and Machine Learning Engineer

Hewlett Packard Enterprise Company in • Spring (TX)

Hybrid
USD 121,000 - 277,000
Hybrid work model
AI and Machine Learning Engineer
AI and Machine Learning Engineer

Hewlett Packard Enterprise • Spring (TX)

On-site
USD 121,000 - 277,000
AI Solution Engineer
AI Solution Engineer

Hewlett Packard Enterprise • Town of Texas (WI)

Remote
USD 152,000 - 349,000
Health & Wellbeing benefits
Personal & Professional Development programs
Inclusive work environment
AI Solution Engineer
AI Solution Engineer

Hewlett Packard Enterprise • United States

Remote
USD 172,000 - 349,000
Comprehensive health benefits
Professional development programs
Senior AI Software Developer
Senior AI Software Developer

Hewlett Packard Enterprise • San Juan (PR)

Hybrid
USD 100,000 - 130,000
Comprehensive benefits for wellbeing
Personal & professional development programs
Job Posting Title HPC Applications & Performance Engineer
Job Posting Title HPC Applications & Performance Engineer

Hewlett Packard Enterprise • Spring (TX)

On-site
USD 105,000 - 243,000
Health & Wellbeing benefits
Personal & Professional Development programs
Inclusive workplace culture
AI & ML Software Engineer
AI & ML Software Engineer

Hewlett Packard Enterprise Development LP • Fort Collins (CO)

Hybrid
USD 144,000 - 273,000
Health & wellbeing
Career development
Inclusive culture
HPC & AI Performance Engineer
HPC & AI Performance Engineer

Hewlett Packard Enterprise • BLOOMINGTON (MN)

On-site
USD 62,000 - 146,000
Comprehensive benefits
Professional development programs
Director, Product Management, AI Support
Director, Product Management, AI Support

Hewlett Packard Enterprise Company in • Cupertino (CA)

Hybrid
USD 194,000 - 413,000
AI Tools Architect / Tech Lead
AI Tools Architect / Tech Lead

Hewlett Packard Enterprise • Fort Collins (CO)

Hybrid
USD 160,000 - 303,000