Senior Software Engineer - ML/CV Infrastructure

Claryo

San Francisco (CA)

On-site

USD 150,000 - 200,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Top-tier medical, dental, and vision coverage
401k with employer matching
Parental leave
Unlimited vacation

Job summary

Claryo is seeking a Staff Software Engineer with a focus on Computer Vision Deployment based in San Francisco. The successful candidate will develop robust infrastructures that power AI-driven warehouse intelligence. Responsibilities include creating and managing distributed cloud GPU infrastructures and building comprehensive computer vision pipelines. Candidates should have over 7 years of experience in software engineering and a proven record in deploying machine learning systems, particularly in production environments. The role is hybrid, requiring 3 days a week in the office.

Qualifications

  • 7+ years of experience in software engineering, specifically in ML infrastructure.
  • Experience deploying computer vision models in real-world environments.
  • Strong programming skills in Python and software engineering practices.

Responsibilities

  • Develop and maintain distributed cloud GPU infrastructure.
  • Build end-to-end computer vision pipelines and integrate them into workflows.
  • Deploy and optimize machine learning models using cloud platforms.

Skills

Machine Learning Infrastructure
Distributed Systems
Cloud Deployment
Python Programming
Computer Vision

Education

B.S. / M.S. in Computer Science, Robotics, or similar

Tools

PyTorch
TensorFlow
Kafka
gRPC
CUDA

Job description

Overview

We\'re looking for a Staff Software Engineer – Computer Vision Deployment to build and scale the infrastructure that powers our AI-driven warehouse intelligence platform. You\'ll own the end-to-end lifecycle of computer vision models — from training pipelines through optimized cloud deployment — ensuring our cutting-edge computer vision and multi-modal AI systems run reliably and efficiently in production. Your work will directly enable the real-time perception and autonomous decision-making capabilities at the core of our platform.

This is a deeply technical role at the intersection of machine learning, distributed systems, and cloud infrastructure. You\'ll design scalable GPU compute clusters, build robust orchestration pipelines, and optimize model serving for low-latency inference at scale. You\'ll work closely with our research scientists, computer vision engineers, and product teams to bridge the gap between experimental models and production-ready systems that operate across diverse warehouse environments. We\'ve found tremendous value in collaborative problem-solving, thus our team works from our SF office three days a week.

Responsibilities
  • Develop and maintain distributed cloud GPU infrastructure for large-scale world model training and low-latency inference.

  • Build end-to-end computer vision pipelines — from data ingestion and preprocessing through model training, evaluation, and deployment — and integrate them into core product workflows.

  • Deploy and optimize state-of-the-art machine learning models in the cloud using model serving platforms and inference optimization techniques, including VLMs and VLAs.

  • Design and operate orchestration systems that enable both engineers and non-engineers to build and manage data and ML pipelines.

  • Establish monitoring, benchmarking, and evaluation frameworks to ensure model performance and reliability in production environments.

Required Experience
  • B.S. / M.S. in Computer Science, Robotics, or similar technical field, or equivalent practical experience.

  • 7+ years of professional software engineering experience, with at least 3 years in machine learning infrastructure — developing, scaling, training, deploying, and optimizing large-scale ML systems from data to model.

  • Track record of deploying computer vision models in production environments with real-world constraints.

  • Experience with distributed messaging and compute systems (Kafka, gRPC, ROS2, or similar).

  • Strong programming skills in Python with solid software engineering practices.

Preferred Experience
  • Experience developing, running, and managing orchestration systems (Flyte, Temporal, Airflow, or similar) for ML and data pipelines.

  • Proficiency with ML frameworks (PyTorch, TensorFlow, DeepSpeed) and model serving platforms (TorchServe, TensorFlow Serving, NVIDIA Triton Inference Server, or similar).

  • Deep understanding of state-of-the-art machine learning models such as auto-regressive transformers and familiarity with inference optimization techniques (TensorRT, quantization, custom kernels).

  • Experience with C++ or CUDA programming for GPU acceleration.

  • Prior experience working at autonomous vehicles or robotics companies.

Equal Opportunity Statement

We’re an equal opportunity employer that values diversity and inclusion. We welcome teammates of all backgrounds and don’t discriminate based on race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or veteran status.

Benefits

At Claryo, we offer a competitive benefits package that supports your health and well-being, including — top-tier medical, dental, and vision coverage, 401k with employer matching, parental leave, and unlimited vacation.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Software Engineer - Computer Vision Deployment
Senior Software Engineer - Computer Vision Deployment

Claryo, Inc. • San Francisco (CA)

On-site
USD 120,000 - 160,000
Top-tier medical coverage
401k with employer matching
Parental leave
+1
Senior Software Engineer - ML Infrastructure
Senior Software Engineer - ML Infrastructure

Claryo • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Medical/Dental/Vision
401k with employer matching
Parental leave
+1
Senior ML Infra Engineer — Real-Time Vision & Cloud GPUs
Senior ML Infra Engineer — Real-Time Vision & Cloud GPUs

Claryo • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Medical/Dental/Vision
401k with employer matching
Parental leave
+1
Computer Vision Engineer
Computer Vision Engineer

Harrison Clarke • San Francisco (CA)

On-site
USD 150,000 - 210,000
Senior Computer Vision/AI Engineer
Senior Computer Vision/AI Engineer

BrightAI Corporation • Palo Alto (CA)

On-site
USD 150,000 - 200,000
CV/ML Platform Engineer
CV/ML Platform Engineer

Allencontrolsystems • Austin (TX)

On-site
USD 90,000 - 120,000
Competitive salary
Health, Dental, Vision Insurance
Paid Time Off
ML Infrastructure Engineer
ML Infrastructure Engineer

Echo • San Francisco (CA)

On-site
USD 120,000 - 160,000
Competitive compensation including stock options
Comprehensive benefits package
401(k) program with matching contributions
Staff Computer Vision Deployment Engineer (Production ML Infra)
Staff Computer Vision Deployment Engineer (Production ML Infra)

Claryo • San Francisco (CA)

On-site
USD 150,000 - 200,000
Top-tier medical, dental, and vision coverage
401k with employer matching
Parental leave
+1
Senior ML Platform Engineer
Senior ML Platform Engineer

Echo • San Francisco (CA)

On-site
Senior Computer Vision Engineer ID72408
Senior Computer Vision Engineer ID72408

AgileEngine, LLC. • Austin (TX)

On-site
USD 130,000 - 180,000
Professional growth
Competitive compensation
Exciting projects
+1