Forward Deployed Machine Learning Engineer

Advatix

Northern (KY)

Hybrid

USD 170,000 - 270,000

Full time

29 hours ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

HRforGrowth is seeking a Forward Deployed Machine Learning Engineer for its US-based team. The role focuses on building AI benchmarks, evaluation systems, and backend infrastructure to support enterprise evaluations across multiple AI domains.

You will own end-to-end projects from feasibility to production, partner with researchers and early customers, and design scalable data pipelines and orchestration for large-scale ML workloads.

Qualifications

  • 4+ years of professional engineering experience with hands-on ML model evaluation.
  • 3–8 years across ML engineering, evaluation systems, benchmarks, and end-to-end ownership.
  • Experience deploying end-to-end ML evaluation or benchmark systems to production for foundation models.
  • Hands-on with ML evaluation frameworks and benchmark design (LLM-as-a-judge).
  • Ownership of backend and infrastructure: storage and orchestration.
  • Experience building benchmarks, evaluations, or data pipelines for LLMs.
  • Experience deploying ML systems in production and customer-facing engineering.
  • Ability to scope ambiguous problems and deliver solutions.

Responsibilities

  • Define, design, and build AI benchmarks and evaluation systems with stakeholders.
  • Develop benchmarks and evaluation frameworks across multiple AI domains.
  • Build and own backend infrastructure supporting AI model evaluation.
  • Design and maintain data pipelines, execution environments, storage, and orchestration.
  • Create sandboxed environments for agentic evaluations with tools and multi-step tasks.
  • Develop and deploy end-to-end evaluation systems for foundation models.
  • Own engineering for customer engagements from requirements to delivery.
  • Collaborate with enterprise customers to translate requirements into solutions.
  • Identify repeatable evaluation patterns for scalable infrastructure and products.
  • Identify infra gaps informing future product development.
  • Scope ambiguous problems and drive implementation through delivery.
  • Collaborate with AI researchers to translate concepts into production systems.
  • Develop scalable systems for large-scale ML evaluation workloads.
  • Communicate status, tradeoffs, and solutions clearly to customers and stakeholders.

Skills

ML evaluation
Backend infrastructure
Customer-facing engineering
Data pipelines
Production deployment

Education

Bachelor's degree in Computer Science, Physics, or related technical field

Tools

Storage
Orchestration
Benchmarks

Job description

Forward Deployed Machine Learning Engineer

Location: , , United States,

Department: Information Technology, Type: Full Time

Role: Forward Deployed Machine Learning Engineer

Our Client is seeking a Forward Deployed Machine Learning Engineer (FDE MLE) with strong experience in machine learning evaluation, benchmark systems, backend infrastructure, and customer-facing engineering. This individual will be the first Machine Learning Engineer dedicated to the organization's Benchmarks and Evaluations vertical and will work closely with the General Manager, researchers, and early customers.

The successful candidate will help establish the technical foundation for evaluating AI models across different domains and modalities. This is a highly hands-on role combining machine learning engineering, backend infrastructure, data pipelines, evaluation systems, and customer-facing technical delivery.

The ideal candidate is comfortable operating in ambiguous, fast-moving environments and can independently take technical problems from feasibility through production delivery. Strong customer engagement skills and experience owning technical projects end-to-end are essential.

Location: Remote
Job Type: Full-Time
Work Setup: Remote
Compensation: $170,000 – $270,000 per year
Visa Sponsorship: Not available; candidates must be authorized to work in the United States without visa sponsorship.

Key Responsibilities
  • Partner directly with the General Manager, researchers, and early customers to define, design, and build AI benchmarks and evaluation systems.
  • Develop benchmarks and evaluation frameworks across multiple AI domains and modalities.
  • Build and own backend infrastructure supporting AI model evaluation.
  • Design and maintain data pipelines, execution environments, storage systems, and orchestration infrastructure.
  • Build sandboxed environments for agentic evaluations involving tools, code execution, and multi-step tasks.
  • Develop and deploy end-to-end evaluation systems for foundation models.
  • Own the engineering component of customer engagements from initial requirements through technical delivery.
  • Work directly with enterprise customers to understand technical requirements and develop practical evaluation solutions.
  • Identify repeatable evaluation patterns that can be transformed into scalable infrastructure and product capabilities.
  • Identify infrastructure gaps and opportunities that can inform future product development.
  • Scope ambiguous technical problems, assess feasibility, and independently drive solutions through implementation and delivery.
  • Collaborate with AI researchers and other technical stakeholders to translate research concepts into reliable production systems.
  • Develop scalable systems capable of supporting large-scale machine learning evaluation and benchmark workloads.
  • Communicate technical concepts, project status, tradeoffs, and solutions clearly to customers and internal stakeholders.
Required Qualifications
  • Minimum 4+ years of professional engineering experience with hands-on machine learning model evaluation experience.
  • Minimum 3–8 years of relevant experience across machine learning engineering, evaluation systems, benchmark development, and end-to-end technical ownership.
  • Demonstrated experience deploying end-to-end ML evaluation or benchmark systems to production for foundation models.
  • Strong hands-on experience with ML evaluation frameworks and benchmark design, including approaches such as LLM-as-a-judge.
  • Prior ownership of backend and infrastructure systems, including:
  • Storage
  • Orchestration
  • Experience building benchmarks, evaluations, or human-data pipelines for large language models (LLMs) is strongly preferred.
  • Experience building data pipelines capable of supporting large-scale workloads.
  • Experience working with or deploying machine learning systems in production.
  • Customer-facing engineering experience, including direct interaction with enterprise customers.
  • Demonstrated ability to independently scope and execute ambiguous technical problems from feasibility through delivery.
  • Strong written communication skills for customer-facing and cross-functional technical work.
  • Strong bias toward action and ability to operate effectively in fast-moving, high-ambiguity environments.
  • Bachelor's degree or higher in Computer Science, Physics, or a related technical field.
Required Experience & Environment

Candidates should demonstrate experience in at least one of the following types of environments:

  • Early-stage B2B startups with demonstrated traction.
  • Forward-deployed engineering organizations or companies with a strong FDE model.
  • High-ownership generalist roles within larger organizations, such as an Office of the CTO or internal machine learning evaluation team.
  • Fast-moving engineering environments where individuals have owned projects from initial concept through production deployment.
Customer-Facing Engineering Expectations

The ideal candidate should be comfortable:

  • Working directly with enterprise customers.
  • Translating customer requirements into technical specifications.
  • Managing technical components of customer engagements.
  • Explaining complex ML evaluation concepts clearly to technical and non-technical stakeholders.
  • Balancing customer-specific requirements with the development of reusable infrastructure.
  • Working under tight customer deadlines while maintaining engineering quality.
  • Taking ownership of technical delivery from initial discovery through deployment.
Preferred Qualifications
  • Experience working directly with AI researchers or foundation model labs.
  • Experience developing evaluation systems for foundation models or LLM-based applications.
  • Published research, papers, or open-source contributions related to ML evaluations or benchmarks.
  • GitHub contributions involving ML evaluation, benchmark systems, LLM evaluation, or related infrastructure.
  • Experience with agentic AI evaluation environments involving tools, code execution, or multi-step workflows.
  • Experience developing human-data pipelines for machine learning evaluation.
  • Experience working across multiple AI modalities or evaluation domains.
Role Expectations
  • High ambiguity tolerance: Comfortable solving problems where requirements and approaches are not fully defined.
  • Bias to action: Able to move quickly from concept and feasibility assessment to implementation.
  • End-to-end ownership: Takes responsibility for technical outcomes rather than isolated engineering tasks.
  • Customer orientation: Comfortable working directly with enterprise customers and incorporating their requirements into technical solutions.
  • Infrastructure depth: Capable of building the backend systems required to support production ML evaluations.
  • Evaluation expertise: Strong understanding of benchmark design, evaluation methodology, and ML evaluation frameworks.
  • Strong communication: Able to document and communicate technical decisions clearly.
  • Generalist mindset: Comfortable moving between ML evaluation, backend infrastructure, data pipelines, and customer-facing technical work.
Backgrounds Less Aligned With This Role
  • This position requires both ML evaluation expertise and strong production engineering/customer-facing experience. Candidates may be less aligned when their background is primarily:
  • Research-oriented MLE work without meaningful production deployment or engineering ownership.
  • Machine learning research without experience building and deploying evaluation infrastructure.
  • Long-term product engineering experience without demonstrated experience in rapid prototyping or highly ambiguous environments.
  • Experience limited exclusively to stealth startups or small B2C companies without relevant B2B, enterprise, or customer-facing engineering experience.
  • ML engineering experience without hands-on benchmark or evaluation system development.
  • Engineering experience without direct customer-facing responsibilities where customer engagement is a significant part of the role.
Equal Employment Opportunity

In its recruiting and search practices, HRforGrowth does not discriminate on the basis of race, color, religion, national origin, age, marital status, physical or mental disability, sex, sexual orientation, gender, or gender identity, and welcomes applications from all qualified individuals. HRforGrowth evaluates and presents candidates to its clients without regard to any protected characteristic. HRforGrowth endeavors to advise the hiring organization to comply with all applicable federal, state, and local equal employment opportunity laws — including those enforced by the U.S. Equal Employment Opportunity Commission (EEOC) — in its hiring decisions, terms of employment, and workplace practices.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Forward Deployed ML Engineer
Forward Deployed ML Engineer

Alexander Chapman • United States

On-site
USD 140,000 - 190,000
Forward Deployed Machine Learning Engineer
Forward Deployed Machine Learning Engineer

raydar • Northern (KY)

Hybrid
USD 170,000 - 270,000
Competitive equity
Forward Deployed Engineer
Forward Deployed Engineer

EVB • United States

Remote
USD 180,000 - 250,000
Equity compensation
Health-insurance premium reimbursement
Paid time off
+1
Machine Learning Engineer / Forward Deployed
Machine Learning Engineer / Forward Deployed

Oscar • California (MO)

On-site
USD 120,000 - 180,000
Equity
Paid Time Off
Medical, dental, and vision coverage
Remote Forward-Deployed ML Engineer — Benchmark & Eval
Remote Forward-Deployed ML Engineer — Benchmark & Eval

Advatix • Northern (KY)

Hybrid
USD 170,000 - 270,000
Forward Deployed Engineer
Forward Deployed Engineer

Eliza Solutions Corp. • New York (NY), Northern (KY)

Hybrid
USD 120,000 - 160,000
Competitive compensation
Equity options
Travel opportunities for on-site work
+2
Forward Deployed Engineer (FDE) - NYC
Forward Deployed Engineer (FDE) - NYC

United States Digital Space LLC • United States

Hybrid
USD 140,000 - 230,000
Relocation assistance
Forward Deployed Engineer (FDE) - SF
Forward Deployed Engineer (FDE) - SF

United States Digital Space LLC • United States

Hybrid
USD 150,000 - 210,000
Forward Deployed Senior Machine Learning / Applied AI Engineer – 25% travel across D.C. and areas across East Coast – Competitive Salary + Equity
Forward Deployed Senior Machine Learning / Applied AI Engineer – 25% travel across D.C. and areas across East Coast – Competitive Salary + Equity

Orbis Group • Washington, Baltimore (MD)

Hybrid
USD 160,000 - 210,000
Equity
Forward Deployed Engineer
Forward Deployed Engineer

Eliza • United States

On-site
USD 90,000 - 120,000
Competitive compensation
Equity options
Flexible work across industries