Senior Performance Engineer, Intel Stack

Sarvam

Bengaluru

On-site

INR 1,500,000 - 2,500,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Sarvam in Bengaluru is seeking an experienced ML Engineer to enhance AI solutions using Intel technologies. Your role involves managing edge models on Intel platforms, optimizing CPU processes, and leading the OpenVINO build process.

The ideal candidate should have over 5 years of experience in ML deployment, including expertise in OpenVINO and Intel inference stacks. Join our dynamic team to shape AI for India's future.

Qualifications

  • 5+ years experience in ML deployment with at least 2 years on Intel inference stacks.
  • Proven production experience with OpenVINO.
  • Knowledge of ONNX Runtime EP.

Responsibilities

  • Manage Sarvam’s edge models on Intel NPU/dGPU/iGPU within SLAs.
  • Own the OpenVINO build/quantization process.
  • Optimize x86 CPU for fallback paths.

Skills

ML deployment
Intel inference stacks
OpenVINO
x86 CPU optimization
AVX-512

Tools

VTune
perf
ONNX Runtime

Job description

About Sarvam

Sarvam is building the bedrock of Sovereign AI for India. The company is developing India’s full-stack sovereign AI platform, building across research, models, infrastructure and applications with a singular focus on making AI genuinely work for India. Sarvam works with leading enterprises and public institutions and is backed by Lightspeed, Peak XV, and Khosla Ventures. Sarvam partners with India’s leading brands, including Tata Capital, SBI Life, CRED, IDFC, and LIC.

About the Role

Own sarvam's Intel surface end-to-end: Intel NPU via OpenVINO (Meteor Lake, Lunar Lake, vPro AI PCs), Intel integrated and discrete GPU via OpenVINO and ONNX Runtime, and x86/AMD64 CPU optimization for fallback paths. You're the technical face to Intel's ecosystem and OpenVINO team.

What You’ll Do
  • Land Sarvam’s edge models on Intel NPU / dGPU / iGPU inside defined SLAs
  • Own the OpenVINO build/quantization recipe for the team's models, including the driver-version compatibility matrix
  • Drive x86 CPU optimization for the universal fallback path - AVX-512, AMX, threading strategy.
  • Own the Intel device-CI pool and regression detection across OpenVINO upgrades.
What We're Looking For
  • 5+ years on ML deployment with 2+ years on Intel inference stacks.
  • Production OpenVINO experience including model conversion, accuracy validation post-quantization, and driver-version pinning.
  • ONNX Runtime EP knowledge - when OpenVINO EP vs. native OpenVINO, when to fall back to CPU EP.
  • x86 CPU profiling and optimization - VTune, perf, comfort reading hot loops at the assembly level when needed.
  • AVX-512 or AMX intrinsics is a strong plus, not a must-have.
Bonus Points
  • Direct prior interaction with the OpenVINO team or Intel ecosystem partners.
  • Custom OpenVINO operator authoring.
Why Sarvam?

Sarvam is a fast-moving, high talent-density team building full-stack AI for India, working on problems that push the frontiers of AI with real population-scale impact.

  • Work alongside researchers, engineers, builders, and business leaders who move fast and hold each other to a very high bar
  • High ownership and high impact, from day one
  • Everything we do is AI-first, from the way we build and ship to the way we think about problems
  • You can work on problems that could change how an entire country learns, works, and communicates

If you want to work on problems at the frontier of AI in India, Sarvam is the place to be.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Performance Engineer, On-Device Inference
Performance Engineer, On-Device Inference

Sarvam • Bengaluru

On-site
INR 1,000,000 - 1,500,000
Principal Forward Deployed Software Engineer (FDSE) - OnDevice AI
Principal Forward Deployed Software Engineer (FDSE) - OnDevice AI

Sarvam • Bengaluru

On-site
INR 2,500,000 - 4,000,000
High ownership and impact
Collaboration with top talent
Fast-paced, AI-first environment
Senior Forward Deployed Software Engineer (FDSE) - OnDevice AI
Senior Forward Deployed Software Engineer (FDSE) - OnDevice AI

Sarvam • Bengaluru

Hybrid
INR 2,500,000 - 3,500,000
Principal Forward Deployed Software Engineer (FDSE) - OnDevice AI
Principal Forward Deployed Software Engineer (FDSE) - OnDevice AI

Sarvam • Bengaluru

On-site
INR 2,500,000 - 3,500,000
Principal Forward Deployed Software Engineer (FDSE) - OnDevice AI
Principal Forward Deployed Software Engineer (FDSE) - OnDevice AI

Sarvam AI • Bengaluru

On-site
INR 2,000,000 - 3,500,000
High ownership
Impact from day one
AI-first work culture
ML Engineer (Training Infra), Foundational Models
ML Engineer (Training Infra), Foundational Models

Sarvam • Bengaluru

On-site
INR 1,200,000 - 2,000,000
High ownership and impact
AI-first approach
Staff Engineer, API Platform
Staff Engineer, API Platform

Sarvam • Bengaluru

On-site
INR 1,200,000 - 1,800,000
Forward Deployed Software Engineer, Model API
Forward Deployed Software Engineer, Model API

Neara • Bengaluru

On-site
INR 1,000,000 - 1,500,000
High ownership and impact
Collaborative team environment
Opportunity to influence India's AI landscape
Product Manager On-Device & Edge AI
Product Manager On-Device & Edge AI

Sarvam AI • Bengaluru

On-site
INR 2,000,000 - 3,000,000
Product Manager On-Device & Edge AI
Product Manager On-Device & Edge AI

Sarvam • Bengaluru

On-site
INR 1,500,000 - 2,500,000
High ownership and impact from day one
Work with a high-talent team
AI-first focused work approach