Foundational Model Engineer — Multimodal & Agentic Medical AI Systems

SAIGroup

Bengaluru

On-site

INR 1,500,000 - 2,500,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive compensation
Opportunity to shape technical architecture
Collaboration with researchers and clinicians

Job summary

A leading AI technology company in Bengaluru is seeking a Foundational Model Engineer to build the technical backend for advanced multimodal and agentic medical AI systems. The role involves designing and optimizing large-scale training pipelines using cutting-edge technologies like PyTorch, JAX, and DeepSpeed. Ideal candidates should have a deep understanding of GPU management and experience with high-performance ML systems. Competitive compensation and the chance to shape groundbreaking AI solutions in healthcare are offered.

Qualifications

  • Strong engineering experience with PyTorch, JAX, or DeepSpeed.
  • Deep understanding of GPU internals and high-performance data pipelines.
  • Experience building large-scale ML pipelines, especially for multimodal workloads.

Responsibilities

  • Architect and maintain large-scale training pipelines for multimodal models.
  • Optimize training performance across A100/H100 clusters.
  • Build and optimize inference runtimes for 3D-aware models.

Skills

PyTorch
JAX
DeepSpeed
CUDA kernels
GPU internals

Tools

Slurm
Kubernetes
Ray

Job description

Foundational Model Engineer — Multimodal & Agentic Medical AI Systems
About the Role

You will be one of the earliest engineering hires responsible for building the technical backbone that powers our 3D-volume foundation model and the agentic medical AI systems built on top of it.

This role blends ML systems engineering,high-performance computing, andfoundation-model infrastructure, enabling our research scientists to train and deploy cutting‑edge multimodal models at scale.

You will design the pipelines, tooling, distributed systems, and evaluation frameworks that make world‑class research possible—and usable in clinical settings.

If you’re the kind of engineer who loves training clusters, PyTorch internals, scalable data loaders, CUDA kernels, model parallelism, and agentic inference systems, this is your role.

What You Will Work On
  • Architect and maintain large‑scale training pipelines for multimodal foundation models (3D volumes + text).
  • Implement distributed training using data parallelism, tensor parallelism, pipeline parallelism, and FSDP/ZeRO strategies.
  • Optimize training performance across A100/H100 clusters, including kernel‑level optimizations and memory efficiency tuning.
  • Build scalable ingestion, preprocessing, and storage systems for 3D medical volumes, DICOM series, voxel grids, and text datasets.
  • Create multimodal data loaders and augmentation pipelines for high‑throughput training.
  • Work on dataset versioning, weak‑label pipelines, and automatic metadata extraction.
Model Serving & Agent Runtime
  • Build and optimize inference runtimes for 3D‑aware models and LLM‑based medical agents.
  • Develop robust APIs and service layers for clinical workflows (retrieval, reporting, case summarization, multi‑step agent chains).
  • Implement caching, quantization, batching, vector search, and agent orchestration.
  • Develop tools for researchers: experiment launchers, logging/visualization dashboards, model evaluation notebooks, and reproducibility tooling.
  • Partner closely with scientists on rapid model iteration, ablations, and experimental design.
  • Participate in internal "ML performance tiger teams" to squeeze maximum throughput from models and data pipelines.
Why This Role Appeals to Top‑Tier ML Systems Engineers
  • You get to build the entire foundational stack behind frontier multimodal models.
  • Rare opportunity to combine 3D infrastructure,LLM agents,medical workflows, and distributed systems.
  • Direct collaboration with researchers working on CLIP‑style models, Chitrarth‑type VLMs, document foundation models, and 3D multimodal architectures.
  • Massive technical scope with freedom to propose new tools, new pipelines, new optimization strategies.
  • Direct impact: your work will enable clinical‑grade AI systems used in radiology and beyond.
What We’re Looking For
  • Strong engineering experience with PyTorch,JAX, orDeepSpeed, plus hands‑on distributed training expertise.
  • Deep understanding of GPU internals, CUDA kernels, NCCL, memory profiling, and high‑performance data pipelines.
  • Experience building large‑scale ML pipelines, especially for multimodal or heavy‑data workloads (video, 3D, imaging).
  • Familiarity with cloud or on‑prem HPC scheduling: Slurm, Kubernetes, Ray, etc.
  • Proficiency in Python + C++/CUDA; strong command of Linux systems.
  • Ability to collaborate deeply with researchers, contribute ideas, and own end‑to‑end engineering projects.
Nice to Have
  • Experience with 3D data (MRI/CT, LiDAR, voxels, meshes, NeRFs).
  • Exposure to vector search (FAISS, Milvus, Annoy) and embedding retrieval systems.
  • Experience with agent frameworks, LLM serving, or multimodal inference pipelines.
  • Contributions to open‑source ML systems or performance optimization libraries.
  • Background in healthcare/medical imaging pipelines (DICOM, PACS, segmentation workflows).
What We Offer
  • Competitive compensation.
  • Opportunity to build the core infrastructure for India’s first 3D multimodal foundation model.
  • Close collaboration with researchers, clinicians, and product teams.
  • Autonomy, ownership, and the chance to shape the technical architecture from the ground up.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Manager — Multimodal Medical Foundation Models
Data Manager — Multimodal Medical Foundation Models

SAIGroup • Bengaluru

On-site
INR 800,000 - 1,500,000
Competitive compensation
Access to unique medical datasets
Collaboration with leading scientists
Senior Agentic AI Engineer — Agentic Medical AI
Senior Agentic AI Engineer — Agentic Medical AI

SAIGroup • Bengaluru

On-site
INR 1,200,000 - 1,800,000
Competitive compensation
Access to high-compute clusters
Cross-functional collaboration
ML Engineer (Training Infra), Foundational Models
ML Engineer (Training Infra), Foundational Models

Sarvam • Bengaluru

On-site
INR 1,200,000 - 2,000,000
High ownership and impact
AI-first approach
Senior Developer — Agentic Clinical Workflow & Orchestration
Senior Developer — Agentic Clinical Workflow & Orchestration

SAIGroup • Bengaluru

On-site
INR 4,000,000 - 7,000,000
Competitive compensation
High-compute resources
Freedom to publish and innovate
Senior ML Backend Engineer
Senior ML Backend Engineer

Jobgether • India

On-site
INR 3,500,000 - 6,500,000
Advanced ML projects
Geospatial analytics exposure
Cloud-native infra
+2
ML Engineer, CloudlyCare
ML Engineer, CloudlyCare

Cloudly Inc • India

On-site
Performance-based commission structure
Two annual festive bonuses
Fully subsidized lunch and snacks
+1
Senior AI/ML Engineer
Senior AI/ML Engineer

SourcingXPress • Bengaluru

On-site
INR 3,000,000 - 5,000,000
Senior AI/ML Engineer/ Developer
Senior AI/ML Engineer/ Developer

RADcube • Hyderabad

On-site
INR 2,500,000 - 3,500,000
Platform Engineer (ML Infrastructure)
Platform Engineer (ML Infrastructure)

TANUH • Bengaluru

On-site
INR 2,500,000 - 4,200,000
AI/ML Engineer
AI/ML Engineer

Jash Data Sciences Pvt. Ltd. • Pune District

On-site
INR 800,000 - 1,200,000
Competitive salary
Learning opportunities
Exposure to latest AI technologies