An AI technology startup is seeking a Benchmarking Specialist in Palo Alto to design and execute ML evaluation benchmarks. You'll work closely with the R&D team to define data standards and maintain documentation. The ideal candidate has experience in ML/LLM evaluation and is fluent in English. This is a full-time position with remote work possibilities, targeting an immediate start date. Competitive compensation will be based on your profile and location.
Qualifications
Experience with ML/LLM evaluation, data science, or technical product roles.
Comfortable reading papers and translating technical details.
Fluent in English and respectful of others.
Responsibilities
Design and execute benchmarks to guide AI model evaluation.
Collaborate with R&D team to build evaluation infrastructure.
Maintain documentation for datasets and benchmarks.
Skills
ML/LLM evaluation
Data science
Technical product roles
Reading papers
High-quality data
Documentation
Job description
An AI technology startup is seeking a Benchmarking Specialist in Palo Alto to design and execute ML evaluation benchmarks. You'll work closely with the R&D team to define data standards and maintain documentation. The ideal candidate has experience in ML/LLM evaluation and is fluent in English. This is a full-time position with remote work possibilities, targeting an immediate start date. Competitive compensation will be based on your profile and location.