AI Field Engineer - Strategic Partnerships

Fireworks AI

San Mateo (CA)

On-site

USD 200,000 - 260,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Equity
Competitive benefits

Job summary

Fireworks AI is hiring an AI Field Engineer for Strategic Partnerships in the United States. You will own technical delivery, co-sell motions, and reference architectures with enterprise partners.

You will build end-to-end POCs, run benchmarks, and drive deployments across Foundry, Azure OpenAI, and ML stacks, balancing customer outcomes with product direction. You will work with partner teams, craft scalable deployment patterns, and translate field signals into roadmap feedback.

Qualifications

  • 3+ years in a pre-sales, partner engineering, forward-deployed, or technical consulting role.
  • Proven ability to build production software with customers, not just advise on it.
  • Strong Python skills; comfortable reading, writing, and debugging production code.
  • Familiarity with LLM inference: latency, throughput, batching, quantization, outputs, function calling.
  • Real experience with fine-tuning — LoRA at minimum; RFT a strong plus.
  • Deep familiarity with the Azure AI stack: Foundry, OpenAI Service, ML, AKS, Entra/RBAC for AI workloads.
  • Exceptional communication: run a sharp discovery call, present to a VP, and debug latency with an ML engineer.

Responsibilities

  • Be the technical lead on co-sell motions with Strategic Partners — joint reference architectures and POCs.
  • Build end-to-end POCs with partner engineering teams inside their codebases and infra.
  • Run load tests and establish latency, throughput, and cost baselines; tune deployments.
  • Deploy and validate new model families on inference frameworks (vLLM, SGLang) and report results.

Skills

Python
Kubernetes
LLM inference
LoRA fine-tuning
Azure Foundry
Azure OpenAI
Azure ML
Strong communication
Partner engineering
Fine-tuning (SFT/RFT)

Tools

Azure Foundry
Azure OpenAI Service
Azure ML

Job description

About Us

Fireworks is the platform for specialized intelligence, enabling companies to build, train, and serve AI models tailored to their own data, workflows, and products. Founded by the team behind PyTorch and backed by AMD, Atreides, Benchmark Capital, Index Ventures, Lightspeed, NVIDIA, Sequoia Capital, and TCV, Fireworks powers production AI with hundreds of state-of-the‑art open models across text, image, embedding, audio, and multimodal workloads. Today, Fireworks is a Series D company valued at $17.5 billion, bringing together an ambitious, collaborative team that's building the future of enterprise AI.

About Us

Fireworks is the platform for specialized intelligence, enabling companies to build, train, and serve AI models tailored to their own data, workflows, and products. Founded by the team behind PyTorch and backed by AMD, Atreides, Benchmark Capital, Index Ventures, Lightspeed, NVIDIA, Sequoia Capital, and TCV, Fireworks powers production AI with hundreds of state-of-the‑art open models across text, image, embedding, audio, and multimodal workloads. Today, Fireworks is a Series D company valued at $17.5 billion, bringing together an ambitious, collaborative team that's building the future of enterprise AI.

In The Last Few Months Alone We Launched Fireworks Training, Partnered With Microsoft Azure Foundry, And Published Research Straight From Our Production Systems. A Few Examples Of What That Looks Like In Practice

  • Frontier RL is cheaper than the mega-cluster narrative suggests: we ran cross‑region rollouts using 98% sparse weight deltas and published what we learned. (blog)
  • Open source agents with frontier advisors: matching frontier performance through training and harness engineering. (blog)
  • The fine‑tuning bottleneck is not the algorithm: integration friction and iteration speed are what actually stall teams; we documented the patterns across dozens of customer engagements. (blog)
The Role

As an AI Field Engineer for Strategic Partnerships, you will be one of the technical owners of Fireworks' most strategic partnership. You’ll work closely with Strategic Partner field teams, Partner‑aligned ISVs, and the SIs that run enterprise AI transformation programs to make Fireworks the default inference and fine‑tuning layer in every Partner AI architecture. The role sits at the intersection of engineering, partner development, and customer delivery. You build reference architectures, run benchmarks, debug production integrations, and co‑develop POCs — all while holding your own in executive‑level conversations about strategy, roadmap, and business outcomes.

You spend most of your time building and enabling. You ship code, run joint POCs with Partner field teams, and architect deployments that span Strategic Partners and Fireworks. But you also lead discovery conversations, align partner stakeholders, and translate field signals into product improvements that compress the feedback loop from partner to roadmap.

The Segment

As a Field Engineer aligned with our Partnerships team you own the technical relationship between Fireworks and the Partner ecosystems, Partner field teams, ISVs building on Strategic Partners, and the SIs that deliver AI transformation programs on Strategic Partners. As an example, the Microsoft partnership is a core go‑to‑market bet: clients like UIPath, Stack Blitz, Motif run via Fireworks on Foundry.. Your job is to scale that pattern across the partner ecosystem. These engagements involve large, multi‑stakeholder organizations, so you will need to navigate both the enterprise buyer (IT, security, compliance) and the builder (ML engineers, platform teams, app developers), while building the trusted‑advisor relationships inside Microsoft's field that multiply your reach.

What You'll Work On
Technical Delivery and Deployment
  • Be the technical lead on co‑sell motions with Strategic Partners — joint reference architectures, partner integration patterns, and shared POCs for strategic accounts.
  • Build end‑to‑end POCs and MVPs alongside partner engineering teams, working inside their codebases, infrastructure, and constraints.
  • Run load tests and establish latency, throughput, and cost baselines against realistic customer traffic profiles, and tune deployments to hit those targets.
  • Deploy and validate new model families on inference frameworks (vLLM, SGLang), determining optimal shapes, quantization configs, and serving patterns across workloads.
Model Strategy and Fine‑Tuning
  • Guide customers on model selection, fine‑tuning strategy (SFT, DPO, RFT), and evaluation methodology.
  • Build and run fine‑tuning pipelines directly with customers, navigating trade‑offs between model families, compute cost, and quality targets.
  • Design and implement evaluation frameworks that measure production‑quality metrics, not just benchmark scores
Product Feedback and Platform Improvement
  • Own the feedback loop — surface partner‑driven product gaps to Fireworks engineering, and translate the roadmap back into partner messaging.
  • Ship external technical content: reference architectures, integration guides, and benchmark posts that make it easy for partners to win deals with us.
  • Track pipeline health; flag risks and opportunities to Field leadership weekly
Minimum Qualifications
What We're Looking For:
  • 3+ years in a pre‑sales, partner engineering, forward‑deployed, or technical consulting role.
  • Demonstrated ability to build production software with customers, not just advise on it. You have shipped code running in someone else's production environment.
  • Strong Python skills. Comfortable reading, writing, and debugging production code. Familiarity with Kubernetes and infrastructure engineering.
  • Hands‑on fluency with LLM inference: latency/throughput tradeoffs, batching strategies, quantization, structured outputs, function calling. You can explain why 50ms p99 matters to an enterprise CTO.
  • Real experience with fine‑tuning — LoRA at minimum, RFT a strong plus. You understand when SFT is enough and when it isn't.
  • Deep familiarity with the Azure AI stack: Azure Foundry, Azure OpenAI Service, Azure ML, AKS, Entra/RBAC for AI workloads. You know where Fireworks fits and where it doesn't.
  • Exceptional communication: able to run a sharp discovery call, present to a VP, and debug a latency issue with an ML engineer in the same afternoon.
Preferred Qualifications
  • 5+ years in technical field or engineering roles where you've owned a technical relationship with a hyperscaler or major SI, not just supported one
  • Experience with inference serving frameworks (vLLM, SGLang, TensorRT‑LLM) and tuning deployments for real workloads.
  • Prior role at a hyperscaler, AI‑native cloud, or inference provider.
  • Deep familiarity with other strategic partner stacks: Platforms for AI workloads, network, and identity integration patterns. You know where Fireworks fits and where it doesn't.
  • Experience with agentic frameworks (LangChain, LlamaIndex, or custom tool‑use pipelines) — you understand how inference latency and reliability shapes agent behavior at scale.
  • Background in model evaluation — you understand why benchmark gaming is rampant and what rigorous evals actually look like.
  • You've written a technical blog post or reference architecture that people actually read.
  • Track record taking GenAI POCs from prototype to production‑scale deployments.
On-Target Expectations (Plus Equity)

$200,000 - $260,000 USD

Total compensation also includes meaningful equity in a fast‑growing startup, along with a competitive salary and comprehensive benefits package. Base salary is determined by a range of factors including individual qualifications, experience, skills, interview performance, market data, and work location.

Fireworks AI is an equal‑opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all innovators.

Why Fireworks?
  • Solve Hard Problems: Tackle challenges at the forefront of AI infrastructure, from low‑latency inference to scalable model serving.
  • Build What’s Next: Work with bleeding‑edge technology that impacts how businesses and developers harness AI globally.
  • Ownership & Impact: Join a fast‑growing, passionate team where your work directly shapes the future of AI—no bureaucracy, just results.
  • Learn from the Best: Collaborate with world‑class engineers and AI researchers who thrive on curiosity and innovation.

Fireworks AI is an equal‑opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all innovators.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Field Engineer - Enterprise
AI Field Engineer - Enterprise

Fireworks AI • United States

On-site
USD 140,000 - 200,000
AI Field Engineer - Enterprise
AI Field Engineer - Enterprise

Fireworks AI • New York (NY)

On-site
USD 180,000 - 240,000
AI Field Engineer - Enterprise
AI Field Engineer - Enterprise

Fireworks AI • San Mateo (CA)

On-site
USD 180,000 - 280,000
AI Field Engineer - AI Natives
AI Field Engineer - AI Natives

Fireworks AI • San Mateo (CA)

On-site
USD 180,000 - 260,000
Member of Technical Staff, Enterprise Foundations
Member of Technical Staff, Enterprise Foundations

Fireworks • New York (NY)

On-site
USD 180,000 - 240,000
AI Deployment Strategist
AI Deployment Strategist

Fireworks • San Mateo (CA)

Hybrid
USD 140,000 - 190,000
Strategic Pursuits Account Executive, AI Native
Strategic Pursuits Account Executive, AI Native

Fireworks • San Francisco (CA)

On-site
USD 300,000 - 340,000
Equity
Stock options
Comprehensive benefits
Strategic Pursuits Account Executive, AI Native
Strategic Pursuits Account Executive, AI Native

Fireworks AI • San Francisco (CA)

On-site
USD 300,000 - 340,000
AI Deployment Strategist
AI Deployment Strategist

Fireworks AI • San Mateo (CA)

Hybrid
USD 140,000 - 230,000
AI Deployment Strategist
AI Deployment Strategist

Fireworks AI • New York (NY)

Hybrid
USD 150,000 - 210,000