Director, AI Platform Engineering

NYC Health + Hospitals

New York (NY)

On-site

USD 150,000 - 200,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A major healthcare provider in New York is seeking a Director of AI Platform Engineering to lead the strategic development of AI infrastructure and ensure compliance with health regulations. The role requires a Master's degree in a related field and extensive experience in MLOps or AI engineering, along with expertise in cloud platforms. This position emphasizes collaboration with cross-functional teams to enhance clinical workflows through AI technologies.

Qualifications

  • Master’s degree in Computer Science, Engineering, Information Systems, or related discipline.
  • 5+ years of experience in MLOps or AI platform engineering.
  • Experience in healthcare or regulated environments.

Responsibilities

  • Provide strategic leadership for AI platform infrastructure.
  • Oversee CI/CD pipelines and governance.
  • Define infrastructure for AI model training and serving.

Skills

Leadership
Collaboration
Technical communication
Stakeholder engagement
Change management

Education

Master's Degree in Computer Science or related field
Bachelor's Degree in relevant discipline

Tools

Azure
Kubernetes
Terraform
Docker
CI/CD platforms

Job description

The Director of AI Platform Engineering provides strategic leadership for the cloud, platform, and deployment infrastructure supporting Artificial Intelligence (AI) across the System. This role ensures AI systems used in clinical workflows operate safely, reliably, securely, and are in compliance with applicable laws and NYC Health + Hospitals rules and regulations. The Director leads platform engineering, cloud architecture, Continuous Integration and Continuous Delivery/Deployment (CI/CD) modernization, and reliability functions ensuring that AI tools enhance clinical excellence and protect patient safety.

Essential Duties and Responsibilities
  • Provides strategic leadership for cloud, platform, and infrastructure engineering, developing and leading multi-year roadmaps, standards, and strategies for the secure and scalable deployment of AI products.
  • Oversees architecture and governance of CI/CD pipelines, infrastructure‑as‑code (Terraform), and Kubernetes/Azure Kubernetes Services (AKS) orchestration to support reliable AI deployment.
  • Defines and oversees the infrastructure for the high‑volume, low‑latency data pipelines, feature stores, and data access layers required for training and real‑time serving of AI models.
  • Establishes enterprise‑wide reliability, and monitoring frameworks to ensure stable, and safe operation of AI systems used by clinicians and care teams.
  • Implements platform controls and audit trails to monitor and ensure Responsible AI practices, model explainability (XAI), and checks for model drift and bias on an ongoing basis.
  • Partners with product management, product development, cybersecurity, Machine Learning Operations (MLOps) engineering, and interoperability teams to ensure AI platform readiness, safe integrations.
  • Leads incident management and root‑cause analysis, to minimize disruptions to clinical workflows and drive reliability improvements.
  • Ensures the AI platform and infrastructure provide the necessary controls, logging, and audit capabilities to meet compliance requirements and support AI safety frameworks.
  • Develops long‑term platform resilience, disaster recovery, and cost optimization strategies to support System‑wide AI expansion.
  • Defines and standardizes the platform's toolchain and Application Programming Interface (API) for the Machine Learning (ML) lifecycle, including model experimentation tracking (e.g., MLflow, ClearML), model registry, and automated testing/validation frameworks.
  • Manages a team of platform and infrastructure engineers.
  • Performs other duties as assigned.
Education and Experience
  • Master's Degree from an accredited college or university in Computer Science, Information Systems or Technology, Cybersecurity, Hospital Administration, Health Care Planning, Business Administration, Mathematics, Engineering or Public Administration; and three (3) years of progressively responsible experience in health care information security, multifaced information technology, health and medical service administration, public administration, or a related discipline with an emphasis on systems programming, systems engineering, software developing, or providing technical support as a specialist; two (2) years of which must have been in a related administrative, managerial or supervisory capacity; or,
  • Bachelor’s Degree from an accredited college or university in disciplines, as listed in “1” above; and five (5) years of progressively responsible experience in health care information security, multifaced information technology, health and medical service administration, public administration, or a related discipline with an emphasis on systems programming, systems engineering, software developing, or providing technical support as a specialist; two (2) years of which must have been in a related administrative, managerial or supervisory capacity.
Minimum Qualifications
  • Master’s degree from an accredited college or university in Computer science, Engineering, Information Systems, or related discipline; and,
  • Five (5) years of experience in Machine Learning Operations (MLOps), Machine Learning (ML) engineering, Artificial Intelligence (AI) platform engineering, or operating production Machine Learning (ML) / Large Language Model (LLM) system; or ten (10) years of experience in Software and Data Engineering.
Certifications Preferred
  • Professional certifications in cloud architecture, ML/AI engineering, or DevOps from leading cloud platforms.
Preferred Knowledge Areas, Skills, Abilities, and other Qualifications
  • Figma, Sketch, Adobe XD, or similar design and prototyping tools.
  • Expertise in Azure architecture, Kubernetes/ Azure Kubernetes Services (AKS), Terraform, Continuous Integration and Continuous Deployment (CI/CD), and automation frameworks.
  • Experience supporting production AI/ML systems or mission‑critical workloads.
  • Knowledge of observability tools, monitoring frameworks, and reliability engineering practices.
  • Understanding of security and compliance standards including Health Insurance Portability and Accountability Act of 1996 (HIPAA) and National Institute of Standards and Technology (NIST).
  • Demonstrated leadership, cross‑functional collaboration, and technical communication skills.
  • Strong stakeholder engagement and change‑management skill.
  • Experience in healthcare, public sector, or other regulated environments.
  • Experience deploying or supporting AI/ML, LLM, or agentic AI systems in production.
  • Familiarity with Site Reliability Engineering (SRE) or platform engineering frameworks.
Experience Using the Following Software and/or Platforms
  • Azure cloud services, Docker, Kubernetes/AKS, Terraform, CI/CD platforms (Azure DevOps, GitHub Actions, Jenkins), monitoring/observability tools (Grafana, Azure Monitor), secrets/IAM security tooling.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Programs Director
AI Programs Director

Priority Ondemand • Knoxville (TN)

On-site
USD 120,000 - 160,000
Senior AI Engineer - IT AI and Data Technology
Senior AI Engineer - IT AI and Data Technology

St. Peter's Health • Town of Montana (WI)

On-site
USD 150,000 - 190,000
AI Engineer - IT AI and Data Technology
AI Engineer - IT AI and Data Technology

Worky • Helena (MT)

On-site
USD 150,000 - 210,000
Sr. Systems Engineer - AI
Sr. Systems Engineer - AI

Kansas Ag Connection • Kansas City (KS)

On-site
USD 140,000 - 190,000
Director of AI & Machine Learning
Director of AI & Machine Learning

Accentuate Staffing • Morrisville (NC)

On-site
USD 120,000 - 150,000
Managing Director-Delivery
Managing Director-Delivery

Anblicks Inc. • Dallas (TX), Northern (KY)

Hybrid
USD 180,000 - 240,000
Director, AI Platform & Reliability Engineering
Director, AI Platform & Reliability Engineering

NYC Health + Hospitals • New York (NY)

On-site
USD 150,000 - 200,000
Cloud Engineer
Cloud Engineer

UCLA Health • Los Angeles (CA)

On-site
USD 128,000 - 299,000
Director / Senior Manager - Cloud Platform Engineering & AI Enablement
Director / Senior Manager - Cloud Platform Engineering & AI Enablement

Anblicks • Dallas (TX)

On-site
USD 150,000 - 190,000
Sr. Systems Engineer - AI
Sr. Systems Engineer - AI

Dairy Farmers of America • Kansas City (KS)

On-site
USD 98,000 - 139,000