Production AI/ML Engineer - Pipelines & Scale

Veritone, Inc.

Irvine (CA)

On-site

USD 175,000 - 200,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Incentive compensation
Health benefits
Retirement benefits
Life insurance
Paid time off
Parental leave and benefits
Other employee perks and benefits

Job summary

Veritone, Inc. is seeking an AI/ML Engineer to build and refactor production-grade ML pipelines and model integration layers.

You will collaborate with product, design, data, and infra teams to embed AI capabilities into applications and maintain scalable features. The role requires 3+ years in AI/ML, strong Python skills, and hands-on experience with LLMs, RAG, and inference services across modern stacks including PyTorch, TensorFlow, Hugging Face, and vector DBs.

Qualifications

  • At least 3 years of professional experience building and deploying software systems with AI/ML models in production environments.
  • Deep proficiency with Python and standard ML libraries (PyTorch, NumPy, Pandas, Scikit-learn, Hugging Face).
  • Hands-on experience with LLMs, RAG architectures, prompt engineering, or traditional ML model pipelines and inference serving.
  • Ability to design and implement robust APIs and backend microservices in Python (Go or Node.js a plus).
  • Professional experience with RDBMSs, NoSQL DBs, and Vector Databases.
  • Demonstrated deployment, monitoring, and scaling of ML workloads and production code.

Responsibilities

  • Process: identify and analyze areas in code, model inference workflows, and data pipelines for optimization, efficiency, and latency improvements.
  • Team: partner with product, design, data, and infrastructure teams to integrate AI/ML capabilities into workflows.
  • On-call: participate in on-call support rotation for production ML services, if necessary.
  • Meetings: share knowledge on AI trends, ask questions, and challenge assumptions.
  • Delivery: help the team meet commitments and accept constructive feedback.
  • Growth: continuously improve technical skills and stay current with AI developments.
  • Quality: write maintainable code with unit/integration tests and evaluation benchmarks.
  • Learning: adapt to new frameworks, algorithms, and stacks quickly.
  • Stack familiarity: Python, PyTorch / TensorFlow, Hugging Face, LangChain / LlamaIndex, vector DBs; Docker, Kubernetes, MLflow / Weights & Biases; AWS SageMaker, GCP Vertex AI; backend & API: Python, Go, Node.js, GraphQL, Elasticsearch, Postgres.

Skills

Python
PyTorch
TensorFlow
Hugging Face
LangChain
LlamaIndex
Vector Databases
Pinecone
Qdrant
pgvector
Go
Node.js
GraphQL
Elasticsearch
PostgreSQL
Docker
Kubernetes
MLflow
Weights & Biases
AWS SageMaker
GCP Vertex AI

Tools

Docker
Kubernetes
MLflow
Weights & Biases
Pinecone
Qdrant
pgvector
SageMaker
Vertex AI

Job description

Veritone, Inc. is seeking an AI/ML Engineer to build and refactor production-grade ML pipelines and model integration layers.

You will collaborate with product, design, data, and infra teams to embed AI capabilities into applications and maintain scalable features. The role requires 3+ years in AI/ML, strong Python skills, and hands-on experience with LLMs, RAG, and inference services across modern stacks including PyTorch, TensorFlow, Hugging Face, and vector DBs.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI/ML Engineer — Production Pipelines & MLOps
AI/ML Engineer — Production Pipelines & MLOps

veritone • Irvine (CA)

On-site
USD 175,000 - 200,000
AI/Machine Learning Engineer
AI/Machine Learning Engineer

Veritone, Inc. • Irvine (CA)

On-site
USD 175,000 - 200,000
Incentive compensation
Health benefits
Retirement benefits
+4
Senior ML Engineer — Production AI Pipelines & Optimization
Senior ML Engineer — Production AI Pipelines & Optimization

Venture Global LNG • Arlington (VA)

On-site
USD 157,000 - 185,000
Production ML Engineer Lead - Pipelines & Inference Scale
Production ML Engineer Lead - Pipelines & Inference Scale

Clera • San Francisco (CA)

On-site
USD 130,000 - 160,000
Remote AI Research Engineer — Build Scalable ML Pipelines
Remote AI Research Engineer — Build Scalable ML Pipelines

Bright-Vision-Technologies • United States

Remote
USD 80,000 - 100,000
Remote AI Data Operations Engineer - Scale ML Pipelines
Remote AI Data Operations Engineer - Scale ML Pipelines

Bright Vision Technologies • Euless (TX), Bedford (TX)

On-site
USD 150,000 - 165,000
AI/ML Engineer — Build Production ML Models & Pipelines
AI/ML Engineer — Build Production ML Models & Pipelines

Latitude • Boston (MA), Northern (KY)

Hybrid
USD 120,000 - 200,000
Medical, dental, and vision insurance
401(k) retirement benefits
Paid time off and holidays
+3
AI/ML Engineer: Production Pipelines & RAG
AI/ML Engineer: Production Pipelines & RAG

Tech Economy • Town of Texas (WI)

Hybrid
USD 79,000 - 95,000
Discretionary bonus
401(k) company contribution
Premium health, dental & vision
+2
Remote AI Data Engineer: Scale Data Pipelines for ML
Remote AI Data Engineer: Scale Data Pipelines for ML

Bright Vision Technologies • Farmington Hills (MI)

On-site
USD 80,000 - 100,000
Production ML Engineer: Scalable Pipelines and Low-Latency
Production ML Engineer: Scalable Pipelines and Low-Latency

Evlo AI • Austin (TX)

On-site
USD 120,000 - 180,000