Machine Learning Engineer, Customer Engineering

Anyscale

San Francisco (CA)

On-site

USD 180,000 - 240,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Anyscale is seeking a Customer Engineer to guide customers through onboarding, adoption, and growth on the Anyscale platform. You will troubleshoot complex issues, coordinate with engineering, and help customers maximize their use of Ray-based ML workloads.

The role emphasizes strong communication, ownership, and collaboration across product and engineering teams in a fast-paced environment.

Qualifications

  • 7+ years of experience in a Machine Learning role in a startup-like environment.
  • Excellent communication and interpersonal skills.
  • Strong ownership mindset and willingness to mentor peers.
  • Experience developing data pipelines for training, fine-tuning, and inference/serving of LLMs.
  • Ability to manage multiple customer needs simultaneously.
  • Experience running distributed ML workloads on major cloud platforms.
  • Willingness to collaborate with product and engineering teams to improve the product experience.

Responsibilities

  • Resolve customer issues and ensure successful adoption of Anyscale platform.
  • Act as a technical advisor and internal champion for key customers.
  • Own customer issues end-to-end from troubleshooting to resolution.
  • Participate in follow-the-sun support to resolve high-priority tickets.
  • Track customer bugs and feature requests to influence prioritization and updates.
  • Contribute to internal tools and documentation with best practices.

Skills

ML experience
Strong communication
Ownership
Data pipelines
Multi-tasking
Mentorship
Customer focus

Tools

AWS/EKS
GCP/GKE
Azure/AKS
Kubernetes
Terraform
Github Actions

Job description

About Anyscale

At Anyscale, we’re on a mission to democratize distributed computing and make it accessible to software developers of all skill levels. We’re commercializing Ray, a popular open-source project that’s creating an ecosystem of libraries for scalable machine learning. Companies like OpenAI, Uber, Spotify, Instacart, Cruise, and many more have Ray in their tech stacks to accelerate the progress of AI applications out into the real world.

With Anyscale, we’re building the best place to run Ray, so that any developer or data scientist can scale an ML application from their laptop to the cluster without needing to be a distributed systems expert.

Proud to be backed by Andreessen Horowitz, NEA, and Addition with $250+ million raised to date.

About the role

The Customer Engineer will play a crucial role in the customers’ post-sale journey – helping them to onboard, adopt, and grow on Anyscale, troubleshooting and resolving open customer tickets, and driving consumption.

Anyscale is an ever-evolving platform and hence will require close coordination with our engineering teams to debug complex issues. This is an exciting role for those who are technically curious and passionate about ML/AI, LLM, vLLM, and the role of AI in next-generation applications. It’s an opportunity to make a significant impact in a collaborative, fast-paced environment while building a new segment in this space.

In this role, you’ll be able to
  • Resolve customer issues and help in their successful adoption of Anyscale platform
  • Be a technical advisor, and internal champion for our key customers
  • Own customer issues end-to-end, from troubleshooting, triaging, escalations, and eventual resolution
  • Participate in our follow-the-sun customer support model to ensure continuity in resolving high-priority tickets
  • Keep track of open customer bugs and feature requests to influence prioritization and provide timely customer updates upon resolution
  • Contribute towards improvement of internal tools and documentation of playbooks, guides, and best practices based on observed patterns
  • Habitually provide feedback and collaborate cross-functionally with product and engineering teams to address customer issues with a focus on improving the product experience
  • Build and maintain strong relationships with technical stakeholders within customer accounts
Qualifications
  • 7+ years of experience in a Machine Learning role in a dynamic, fast-paced, startup-like environment
  • Strong organizational skills and ability to manage multiple customer needs simultaneously
  • Proficient at developing data pipelines for training, fine-tuning, and inference/serving of LLMs
  • Experience running and optimizing infrastructure for distributed ML workloads on the major cloud platforms (AWS/EKS, GCP/GKE, or Azure/AKS)
  • Excellent communication and interpersonal skills.
  • Strong sense of ownership, self-motivation, and eagerness to acquire new skills and do new things
  • Willingness to up-level the knowledge and skills of peers through mentorship, trainings, and shadowing
Bonus
  • Experience with Ray
  • Knowledge of MLOps platforms
  • Knowledge of container orchestration platforms (e.g., Kubernetes), infrastructure as code (e.g., Terraform), CI/CD tools (e.g., Github Actions)

Anyscale Inc. is an Equal Opportunity Employer. Candidates are evaluated without regard to age, race, color, religion, sex, disability, national origin, sexual orientation, veteran status, or any other characteristic protected by federal or state law.

Anyscale Inc. is an E-Verify company and you may review the Notice of E-Verify Participation and the Right to Work posters in English and Spanish.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Distributed LLM Inference Engineer
Distributed LLM Inference Engineer

Cerebras • Palo Alto (CA)

On-site
USD 120,000 - 160,000
Stock Options
Healthcare plans with 99% premium coverage
401k Retirement Plan
+6
Distributed LLM Inference Engineer
Distributed LLM Inference Engineer

Anyscale • San Francisco (CA)

On-site
USD 120,000 - 160,000
Stock Options
Healthcare plans
401k Retirement Plan
+6
Software Engineer, Infrastructure
Software Engineer, Infrastructure

Socket.dev • Palo Alto (CA)

On-site
USD 150,000 - 210,000
Senior Site Reliability Engineer, Platform Infrastructure (Foundations)
Senior Site Reliability Engineer, Platform Infrastructure (Foundations)

Anyscale • San Francisco (CA)

On-site
USD 130,000 - 180,000
Forward Deployed Engineer - AI/ML Platforms
Forward Deployed Engineer - AI/ML Platforms

anyscale • San Francisco (CA)

On-site
USD 120,000 - 160,000
Senior Site Reliability Engineer, Platform Infrastructure (Foundations)
Senior Site Reliability Engineer, Platform Infrastructure (Foundations)

Cerebras • Palo Alto (CA)

On-site
USD 130,000 - 170,000
Staff Software Engineer, Platform Infrastructure (Foundations)
Staff Software Engineer, Platform Infrastructure (Foundations)

Anyscale, Inc. • San Francisco (CA)

On-site
USD 180,000 - 280,000
Head of Platform Infrastructure (Foundations)
Head of Platform Infrastructure (Foundations)

Cerebras • San Francisco (CA)

On-site
USD 260,000 - 380,000
Software Engineer, Infrastructure
Software Engineer, Infrastructure

anyscale • United States

On-site
USD 140,000 - 210,000
Forward Deployed Engineer
Forward Deployed Engineer

Anyscale, Inc. • San Francisco (CA)

On-site
USD 140,000 - 190,000