Low-Latency AI Platform Engineer

Cartesia AI, Inc.

San Francisco (CA)

On-site

USD 140,000 - 190,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health Insurance
401(k)
Commuter Allowance
Flexible PTO

Job summary

Cartesia AI, Inc. in San Francisco is seeking a Software Engineer, Platform to advance our mission of real-time multimodal intelligence.

You’ll design and build a scalable inference and serving stack for our SSM foundation models and partner with research and product teams to translate cutting-edge research into impactful products. The role offers significant autonomy to shape our platforms, with exposure to large-scale distributed systems, high performance requirements, and a culture that

Qualifications

  • Strong engineering skills, comfortable navigating complex codebases and monorepos.
  • An eye for craft and writing clean and maintainable code.
  • Diving into new technologies and adapting your stack (Go, Python, Next.js).
  • Experience building large-scale distributed systems with performance, reliability, and observability.
  • Technical leadership with the ability to execute and deliver in ambiguity.
  • Bonus: background in machine learning or generative models.

Responsibilities

  • Design and build low latency, scalable inference and serving stack for our SSM foundation models.
  • Collaborate with research and product teams to translate research into products.
  • Build high-quality data processing and evaluation infrastructure for model training.
  • Enjoy significant autonomy to shape our products and influence AI deployment across devices and applications.

Skills

Distributed systems
Technical leadership
Clean code
Adaptability
ML/generative models (bonus)

Tools

Go
Python
Next.js
Monorepos

Job description

Cartesia AI, Inc. in San Francisco is seeking a Software Engineer, Platform to advance our mission of real-time multimodal intelligence.

You’ll design and build a scalable inference and serving stack for our SSM foundation models and partner with research and product teams to translate cutting-edge research into impactful products. The role offers significant autonomy to shape our platforms, with exposure to large-scale distributed systems, high performance requirements, and a culture that

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Software Engineer – Real-Time Multimodal AI
Software Engineer – Real-Time Multimodal AI

Cartesia • United States

On-site
USD 120,000 - 180,000
Competitive base salary
Equity package
Fully covered health insurance
+6
Software Engineer, Product
Software Engineer, Product

Cartesia • San Francisco (CA)

On-site
USD 120,000 - 150,000
Software Engineer, AI & Developer Acceleration
Software Engineer, AI & Developer Acceleration

Cartesia • San Francisco (CA)

On-site
USD 120,000 - 160,000
Staff Software Engineer, AI Inference Platform
Staff Software Engineer, AI Inference Platform

Cerebras • Sunnyvale (CA)

On-site
USD 140,000 - 200,000
Software Engineer, Platform (India)
Software Engineer, Platform (India)

Cartesia • United States

On-site
USD 120,000 - 180,000
Competitive base salary
Equity package
Fully covered health insurance
+6
Senior Model Serving Engineer - Low-Latency AI Platform
Senior Model Serving Engineer - Low-Latency AI Platform

Menlo Ventures • San Francisco (CA)

On-site
USD 192,000 - 260,000
Comprehensive benefits
Eligibility for annual performance bonus
Equity opportunities
Software Engineer
Software Engineer

Cartesia AI, Inc. • San Francisco (CA)

On-site
USD 140,000 - 190,000
Health Insurance
401(k)
Commuter Allowance
+1
Senior AI Model Serving Engineer — Low-Latency Inference
Senior AI Model Serving Engineer — Low-Latency Inference

Menlo Ventures • San Francisco (CA)

On-site
USD 166,000 - 225,000
Annual performance bonus
Equity options
Comprehensive benefits package
Hands-on AI Platform Support Engineer
Hands-on AI Platform Support Engineer

cartesia • San Francisco (CA)

On-site
USD 80,000 - 110,000
Competitive base salary
Fully covered health insurance
Flexible PTO
+2
Senior ML Infra Engineer: Low-Latency Inference
Senior ML Infra Engineer: Low-Latency Inference

Doist • New York (NY)

Hybrid
USD 170,000 - 250,000
Healthcare
401k plan with matching
Hybrid work model
+1