SpreeAI is seeking a Principal Engineer in San Francisco to build AI infrastructure that powers real-time virtual try-on experiences. The role involves engineering ML platforms and deployment pipelines, with responsibilities including optimizing GPU utilization and establishing observability systems. Candidates should have 10+ years of engineering experience with deep knowledge in ML infrastructure, cloud services, and performance optimization. Join a fast-growing AI company focused on revolutionizing online shopping with cutting-edge technology.
Qualifications
10+ years of software engineering / infrastructure experience, with 5+ years in ML infrastructure or AI platform engineering.
Deep experience with Python, PyTorch, Kubernetes, Docker, cloud infrastructure, and GPU workloads.
Strong understanding of distributed systems and large-scale ML infrastructure design.
Responsibilities
Build and operate SPREEAI’s end-to-end ML platform spanning training, evaluation, deployment, and monitoring.
Enable reliable and scalable inference deployments through standardized serving, orchestration, and monitoring frameworks.
Partner with research teams to productionize new capabilities by providing robust infrastructure.
Skills
Software engineering / infrastructure experience
ML infrastructure
Python
PyTorch
Kubernetes
Docker
Cloud infrastructure
GPU-based workloads
Distributed systems understanding
Inference optimization techniques
Job description
SpreeAI is seeking a Principal Engineer in San Francisco to build AI infrastructure that powers real-time virtual try-on experiences. The role involves engineering ML platforms and deployment pipelines, with responsibilities including optimizing GPU utilization and establishing observability systems. Candidates should have 10+ years of engineering experience with deep knowledge in ML infrastructure, cloud services, and performance optimization. Join a fast-growing AI company focused on revolutionizing online shopping with cutting-edge technology.