AI Modeling Engineer: Voice & Multimodal (Stock Options)

Palona AI

Los Altos (CA)

On-site

USD 140,000 - 210,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Stock options
Benefits: medical/dental/vision/ret/le
Family leave
Disability protection
Paid time off & holidays
Learning and development

Job summary

Palona AI is seeking an applied AI Modeling Engineer to enhance the intelligence, safety, latency, and cost of Palona’s voice and multimodal agents in production environments. You will own model selection, routing, prompting, fine-tuning when justified, and evaluation methodologies, collaborating with product, engineering, and customer teams to move from hypothesis to reliable deployment.

You will build representative datasets, evaluate open models, and translate research into concrete product

Qualifications

  • 3+ years of industrial experience in AI/ML or related technical domain.
  • Strong Python and ML tooling foundations.
  • Experience with production AI systems and model evaluation.

Responsibilities

  • Develop modeling and experimentation strategies for voice, language, and multimodal agents.
  • Build offline and online evaluations measuring accuracy, latency, cost, and user experience.
  • Create datasets from simulations, annotation, and feedback while protecting sensitive data.
  • Evaluate frontier/open-source models and decide build/buy/prompt/fine-tune decisions.
  • Improve prompting, context construction, memory, tool-use policies, and model routing.
  • Design and implement post-training methods when advantageous.
  • Collaborate with speech/real-time engineers to reduce latency and improve ASR/TTS.

Skills

Industry experience
Python
ML foundations
LLMs
Experiment design
Evaluation pipelines

Tools

PyTorch
JAX
Hugging Face

Job description

Palona AI is seeking an applied AI Modeling Engineer to enhance the intelligence, safety, latency, and cost of Palona’s voice and multimodal agents in production environments. You will own model selection, routing, prompting, fine-tuning when justified, and evaluation methodologies, collaborating with product, engineering, and customer teams to move from hypothesis to reliable deployment.

You will build representative datasets, evaluate open models, and translate research into concrete product

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Modeling Engineer
AI Modeling Engineer

Palona AI • Los Altos (CA)

On-site
USD 140,000 - 210,000
Stock options
Benefits: medical/dental/vision/ret/le
Family leave
+3
AI Research Engineer: Vision & VLMs (Stock Options)
AI Research Engineer: Vision & VLMs (Stock Options)

Palona AI • Los Altos (CA)

On-site
USD 150,000 - 210,000
Competitive salary
Stock options
Green card sponsorship
+3
Senior AI Engineer — Real-Time Multimodal Systems
Senior AI Engineer — Real-Time Multimodal Systems

Propio Language Services • Overland Park (KS)

On-site
USD 120,000 - 210,000
Product Engineer, AI Voice Experiences
Product Engineer, AI Voice Experiences

Cartesia AI, Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 210,000
Compensation Competitive base salary
Health Insurance
Parental Leave 9 weeks paternity & 12 
+5
Senior ML Engineer – Voice AI, Production Systems
Senior ML Engineer – Voice AI, Production Systems

Modulate • Somerville (MA)

Hybrid
USD 170,000 - 200,000
Competitive salary + equity
Health benefits
Hybrid work model
+2
Senior PM, Model APIs & Multimodal Dev Experience
Senior PM, Model APIs & Multimodal Dev Experience

Together AI • San Francisco (CA)

On-site
USD 200,000 - 280,000
Equity
Health insurance
Startup benefits
Multimodal AI Systems Architect (AI Engineering)
Multimodal AI Systems Architect (AI Engineering)

Hyphen Connect Limited • San Francisco (CA)

On-site
USD 120,000 - 160,000
Multimodal AI Engineer: Image/Video Generation & Systems
Multimodal AI Engineer: Image/Video Generation & Systems

xAI • Palo Alto (CA)

On-site
USD 180,000 - 440,000
Comprehensive medical, vision, and dental coverage
401(k) retirement plan
Short & long-term disability insurance
+3
Staff Research Engineer, Multimodal Generative AI
Staff Research Engineer, Multimodal Generative AI

synthesia • United States

On-site
USD 180,000 - 260,000
Technical Program Manager, Multimodal
Technical Program Manager, Multimodal

OpenAI • San Francisco (CA)

Hybrid
USD 207,000 - 445,000
Relocation assistance