Senior Developer

ICE

Atlanta (GA)

On-site

USD 180,000 - 240,000

Full time

5 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Intercontinental Exchange, Inc. (ICE) is hiring a full-time AI Platform Developer to architect and manage an enterprise-wide platform for AI model training, deployment, and inference at scale.

You will lead as a Senior Developer within the AI Center of Excellence, guiding MLOps and platform engineering teams and collaborating with data scientists and business stakeholders in Atlanta. The role requires deep expertise in AI/ML training pipelines, model serving, and real-time inference, with strong

Qualifications

  • Extensive experience designing AI/ML training and inference platforms on cloud infrastructure.
  • Deep expertise in model serving and deployment pipelines (TensorFlow Serving, TorchServe, Kubeflow).
  • Proficiency with GPU clusters, distributed training, and performance optimization.
  • Experience building and leading MLOps pipelines, model versioning, and CI for ML models.

Responsibilities

  • Architect and manage enterprise-wide AI training and inference platform infrastructure.
  • Drive innovation, operations excellence, and scalability for AI model serving environments.
  • Provide technical leadership for AI platform development and cross-functional alignment.

Skills

AI platform
MLOps leadership
Kubernetes
Python
Distributed training
GPU clusters
Model serving

Education

PhD or MS/BS in CS/ML/Data Eng

Tools

TensorFlow Serving
TorchServe
MLflow
Kubeflow
Docker
Kubernetes
PyTorch

Job description

Job Purpose

Intercontinental Exchange, Inc. (ICE) presents an opportunity for a full-time AI Platform Developer to join a team responsible for architecting and managing the enterprise-wide platform for AI model training, deployment, and inference at scale. The candidate will serve as a Senior Developer within the AI Center of Excellence team, playing a pivotal role in advancing the firm’s strategic initiative to integrate Generative AI technologies responsibly and sustainably across the enterprise through robust training and inference infrastructure.

Overview
Job Purpose

Intercontinental Exchange, Inc. (ICE) presents an opportunity for a full-time AI Platform Developer to join a team responsible for architecting and managing the enterprise-wide platform for AI model training, deployment, and inference at scale. The candidate will serve as a Senior Developer within the AI Center of Excellence team, playing a pivotal role in advancing the firm’s strategic initiative to integrate Generative AI technologies responsibly and sustainably across the enterprise through robust training and inference infrastructure.

The ideal candidate must possess deep expertise in AI/ML training pipeline architecture, inference optimization, and production model serving platforms leveraging the latest advancements in Generative AI, distributed computing, GPU clusters, model optimization techniques, and high-performance inference systems.

This position demands advanced technical proficiency in training orchestration, model deployment pipelines, inference scaling, and performance optimization, innovative problem-solving capabilities, strong leadership qualities, and the ability to mentor and guide MLOps and platform engineering teams effectively. The role requires strategic vision for AI training and inference infrastructure roadmaps, including compute resource management, model lifecycle optimization, and real-time serving architectures. Exceptional professionalism, proactive collaboration, and outstanding communication skills are essential.

The candidate will actively engage and influence diverse stakeholders across the organization to align training and inference platform capabilities with AI model requirements and business SLAs, ensuring efficient resource utilization, optimal model performance, and cost-effective scaling. Strong written and verbal communication skills are imperative, given the candidate’s responsibility to articulate training efficiency metrics, inference latency optimizations, resource allocation strategies, and platform ROI clearly and persuasively to both technical teams and executive audiences, including presenting model performance benchmarks, infrastructure cost optimization, and platform scalability roadmaps to senior leadership.

Responsibilities
  • Architecting, implementing, and managing enterprise-wide AI inference and training platform infrastructure.
  • Driving innovation, operational excellence, and scalability within AI/ML model serving and training environments.
  • Leading technical strategy for AI platform development and optimization across the organization.
Knowledge And Experience
  • Extensive experience and demonstrated leadership in designing and managing AI/ML training and inference platforms using cloud infrastructure (AWS, Azure, GCP).
  • Deep expertise in ML model serving frameworks (e.g., TensorFlow Serving, TorchServe, MLflow, Kubeflow).
  • Proficiency with GPU cluster management, distributed training and model optimization techniques.
  • Strong experience with AI/ML orchestration platforms, particularly Kubernetes for ML workloads and container technologies including Docker.
  • Comprehensive knowledge of MLOps pipelines, model versioning, A/B testing frameworks, and continuous integration for ML models.
  • Advanced programming skills in Python, experience with ML frameworks (TensorFlow, PyTorch, Hugging Face), and proficiency in performance optimization.
  • Experience with high-performance computing, inference optimization, and real-time model serving architectures.
  • Exceptional problem-solving skills in AI infrastructure challenges and strategic thinking for platform scalability.
  • Proven leadership abilities in guiding cross-functional AI/ML engineering teams and mentoring MLOps engineers.
  • Excellent written and verbal communication skills for technical and executive audiences.
  • Ability to effectively collaborate with data scientists, ML engineers, and business stakeholders to align AI platform capabilities with strategic objectives.
Preferred Knowledge And Experience
  • Advanced degree (PhD with few years' experience, or MS/BS with extensive experience) in Computer Science, Machine Learning, Data Engineering, or related field with focus on AI/ML systems.
  • Strong programming skills in Python with deep knowledge of ML libraries (scikit-learn, TensorFlow, PyTorch, Transformers).
  • Proficiency in ML model deployment frameworks, inference engines, and real-time serving APIs.
  • Working knowledge of vector databases, model registries, and feature stores (e.g., Feast, Tecton).
  • Experience with distributed computing frameworks (Spark, Ray) and GPU programming (CUDA) is highly beneficial.
  • Experience with AI model monitoring, performance tracking, and observability tools (Prometheus, Grafana, MLflow).
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Developer
Senior Developer

ICE Clear Europe Limited • Atlanta (GA)

On-site
USD 150,000 - 210,000
Senior Developer
Senior Developer

ICE • Jacksonville (FL)

On-site
USD 120,000 - 190,000
Senior Developer
Senior Developer

ICE Clear Europe Limited • Georgia

On-site
USD 120,000 - 180,000
Lead AI/ML Platform Engineer
Lead AI/ML Platform Engineer

Staff Financial Group • Atlanta (GA)

Hybrid
USD 119,000 - 180,000
Sr. AI Platform Engineer
Sr. AI Platform Engineer

TechWize • New York (NY)

On-site
USD 120,000 - 160,000
AI/ML Platform Engineer
AI/ML Platform Engineer

Planet Pharma • Indianapolis (IN)

On-site
USD 130,000 - 170,000
Senior AI Platform Engineer — Training & Inference Leader
Senior AI Platform Engineer — Training & Inference Leader

ICE • Atlanta (GA)

On-site
USD 180,000 - 240,000
AI/ML Engineer
AI/ML Engineer

Winaxis LLC • Dallas (TX)

On-site
USD 120,000 - 180,000
AI Platform Engineer
AI Platform Engineer

Accord Technologies Inc • Dallas (TX)

Hybrid
USD 130,000 - 180,000
AI/ML Platform Engineer
AI/ML Platform Engineer

Surge IT • Alexandria (VA)

On-site
USD 120,000 - 150,000