AI Engineer

LiveX AI Inc.

Palo Alto (CA)

On-site

USD 150,000 - 210,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Benefits package
Career growth opportunities

Job summary

LiveX AI Inc. is building next‑generation realtime, interactive AI avatars that see, listen, and respond with natural speech and expressive motion. We seek a strong ML Engineer to train, fine‑tune, and serve generative models powering these avatars.

You will work across the full model lifecycle—from data curation and training of diffusion, video, and multimodal models, to low‑latency inference optimization for live streaming deployments.

Qualifications

  • 3+ years of hands-on ML engineering experience shipping deep learning models.
  • Strong expertise in PyTorch and modern ML workflows.
  • Experience with diffusion, video generation, or multimodal AI is highly desirable.
  • Experience optimizing models for real-time, low-latency inference.

Responsibilities

  • Train and fine‑tune state‑of‑the‑art generative models for talking-head synthesis and avatar video.
  • Design data pipelines for large-scale video, audio, and multimodal datasets.
  • Optimize models for sub-200ms end‑to‑end inference and streaming pipelines.
  • Build and maintain training/evaluation/deployment infra on GPU clusters.
  • Define and track quality metrics like lip-sync accuracy and latency.
  • Collaborate with backend, product, and research teams to ship features to production.
  • Stay current with research in generative video and neural rendering.

Skills

PyTorch
Diffusion models
Video generation
Multimodal AI
Streaming inference
Python
Distributed training

Tools

CUDA
TensorRT
ONNX
DeepSpeed

Job description

Job Overview: LiveX AI is building the next generation of realtime, interactive AI avatars—lifelike digital humans that see, listen, and respond with natural speech and expressive motion. We are hiring a strong Machine Learning Engineer to help us train, fine‑tune, and serve the generative models that power these avatars. You will work across the full model lifecycle—from data curation and training of diffusion, video, and multimodal models, to low‑latency inference optimization for live, streaming deployments.

Your work will directly shape how millions of end users experience AI—turning static chatbots into engaging, face‑to‑face digital humans that deliver VIP‑level customer experiences in real time.

Responsibilities:
  • Train and fine‑tune state‑of‑the‑art generative models for talking‑head synthesis, audio‑driven facial animation, and full‑body avatar video (e.g., diffusion models, DiT, 3D Gaussian Splatting, GAN‑based and transformer‑based video generators).
  • Design data pipelines for large‑scale video, audio, and multimodal datasets—including collection, cleaning, labeling, alignment, and augmentation.
  • Optimize models for realtime, low‑latency serving (sub‑200ms end‑to‑end): model distillation, quantization, pruning, caching, streaming inference, and custom CUDA/Triton kernels where needed.
  • Build and maintain the training, evaluation, and deployment infrastructure on GPU clusters (PyTorch, DeepSpeed, FSDP, Ray, Kubernetes).
  • Define and track quality metrics—lip‑sync accuracy, identity preservation, expression naturalness, latency, and throughput—and drive measurable improvements.
  • Collaborate closely with backend, product, and research teams to ship avatar features into production and iterate rapidly based on real user feedback.
  • Stay current with the latest research in generative video, multimodal AI, and neural rendering, and bring novel techniques into our stack.
Qualifications:
  • 3+ years of hands‑on ML engineering experience (or strong research experience with production‑quality code), with a track record of shipping deep learning models.
  • Deep expertise in PyTorch and modern deep learning workflows; comfortable reading, reproducing, and extending research papers.
  • Strong background in one or more of: diffusion models, video generation, generative AI, computer vision, multimodal AI (audio + vision), speech synthesis, or neural rendering.
  • Experience with talking‑head/audio‑driven animation/lip‑sync/digital human systems, or closely related areas (3D Gaussian Splatting, NeRF, face reenactment, avatar synthesis) is a strong plus.
  • Proven ability to optimize models for realtime inference—familiarity with TensorRT, ONNX, Triton Inference Server, CUDA, mixed‑precision, and streaming pipelines.
  • Solid software engineering fundamentals in Python; experience with distributed training on multi‑GPU/multi‑node setups.
  • Comfortable working in ambiguous and rapidly evolving environments, with a proactive, ownership‑driven mindset.
  • Fast learner with a startup mentality, eager to push the frontier of what interactive AI avatars can do.
  • Publications at top venues (NeurIPS, CVPR, ICCV, SIGGRAPH, ICLR, ICML) or notable open‑source contributions are a plus.
What We Offer:
  • A rare opportunity to build realtime interactive avatars at the intersection of generative video, speech, and multimodal AI.
  • Access to substantial GPU compute and a team that ships to real customers.
  • A collaborative, innovative, and dynamic work environment.
  • Competitive salary, equity, and benefits package.
  • Career growth opportunities in a rapidly expanding sector.

LiveX AI is an Equal Opportunity Employer.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Realtime Generative AI Avatar Engineer
Realtime Generative AI Avatar Engineer

LiveX AI Inc. • Palo Alto (CA)

On-site
USD 150,000 - 210,000
Equity
Benefits package
Career growth opportunities
Member of Technical Staff - Imagine Model
Member of Technical Staff - Imagine Model

xAI • Palo Alto (CA)

On-site
USD 180,000 - 440,000
Comprehensive medical, vision, and dental coverage
401(k) retirement plan
Short & long-term disability insurance
+3
Machine Learning Engineer
Machine Learning Engineer

247Hire • Orlando (FL)

On-site
USD 120,000 - 160,000
Member of Technical Staff - Imagine Model
Member of Technical Staff - Imagine Model

xAI • Seattle (WA)

On-site
USD 180,000 - 440,000
Equity
Comprehensive medical, vision, and dental coverage
401(k) retirement plan
+2
Senior Machine Learning Engineer, 3D Generative AI
Senior Machine Learning Engineer, 3D Generative AI

Alexander Chapman • Los Angeles (CA)

Hybrid
USD 240,000 - 340,000
Comprehensive health/wellness
Unlimited PTO
Professional development support
ML Engineer - Data
ML Engineer - Data

Nuance Labs • Seattle (WA)

On-site
USD 90,000 - 120,000
Conversational Modelling Research Engineer
Conversational Modelling Research Engineer

Tavus • San Francisco (CA)

Hybrid
USD 120,000 - 230,000
Research Scientist - Video Diffusion
Research Scientist - Video Diffusion

Nuance Labs • Seattle (WA)

On-site
USD 150,000 - 210,000
Builder - Senior Software Engineer, AI
Builder - Senior Software Engineer, AI

CV in • Northern (KY)

Hybrid
USD 140,000 - 200,000
Senior/Staff Fullstack Engineer, Avatars
Senior/Staff Fullstack Engineer, Avatars

synthesia • United States

On-site
USD 140,000 - 200,000
25 days leave
Local holidays