AI Compiler Engineer — Real-Time Inference & Optimization

Black Sesame

San Jose (CA)

On-site

USD 120,000 - 170,000

Full time

31 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Black Sesame in San Jose, CA is seeking an experienced Software Engineer for AI Compiler responsibilities, focusing on deploying AI inference and framework integration on BST Intelligence processors. You will work on deep learning infrastructure, data pipelines, and real-time inference, with emphasis on low-precision inference, quantization, and performance optimization.

This role requires strong C++/C or Python, plus ability to define goals, work independently, and stay current with research in

Qualifications

  • Master's or PhD in CS/CE/Applied Math or related field.
  • Ability to define project goals and lead development.
  • Strong C++/C or Python programming and software design, including debugging and performance analysis.
  • Expertise in Cadence DSP programming.
  • Nice to have: experience with DL frameworks like TensorFlow, PyTorch, MXNet.

Responsibilities

  • Contributing to deep learning infrastructure, data pipelines, tools and workflows that lay the foundation for building AI at scale.
  • Writing software to deploy AI models and pipelines in real time applications (inference).
  • Apply low precision inference, quantization, and compression of DNNs.
  • Continuously improve inference latency, accuracy and power consumption of DNNs.
  • Stay up to date with the latest research and innovations in deep learning, implement and experiment with new ideas.

Skills

C++
Python
Software design
Debugging
Performance analysis
Cadence DSP

Education

MS or PhD in CS/CE/Applied Math

Job description

Black Sesame in San Jose, CA is seeking an experienced Software Engineer for AI Compiler responsibilities, focusing on deploying AI inference and framework integration on BST Intelligence processors. You will work on deep learning infrastructure, data pipelines, and real-time inference, with emphasis on low-precision inference, quantization, and performance optimization.

This role requires strong C++/C or Python, plus ability to define goals, work independently, and stay current with research in

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Framework & Inference Engineer
AI Framework & Inference Engineer

Black Sesame • San Jose (CA), Northern (KY)

Hybrid
USD 150,000 - 190,000
Software Engineer, AI Compiler
Software Engineer, AI Compiler

Black Sesame • San Jose (CA)

On-site
USD 120,000 - 170,000
AI Compiler Engineer: ML Accelerator Optimizer
AI Compiler Engineer: ML Accelerator Optimizer

Black Sesame Technologies Inc • San Jose (CA)

On-site
USD 100,000 - 130,000
Software Engineer, AI Framework
Software Engineer, AI Framework

Black Sesame • San Jose (CA), Northern (KY)

Hybrid
USD 150,000 - 190,000
Tech Lead: AI Compiler & Performance Architect
Tech Lead: AI Compiler & Performance Architect

Black Sesame Technologies Inc • San Jose (CA)

On-site
USD 180,000 - 260,000
AI Compiler Research Intern (MLIR, GPU)
AI Compiler Research Intern (MLIR, GPU)

ByteDance • San Jose (CA)

On-site
USD 100,000 - 167,000
ML Accelerator Compiler Engineer
ML Accelerator Compiler Engineer

Black Sesame • San Jose (CA), Northern (KY)

Hybrid
USD 150,000 - 190,000
ML Compiler Engineer for AI Inference Accelerators
ML Compiler Engineer for AI Inference Accelerators

Amazon Inc. • Cupertino (CA)

On-site
USD 165,000 - 224,000
Health insurance
401(k) matching
Paid time off
+1
Senior MLIR Compiler Engineer for AI Silicon
Senior MLIR Compiler Engineer for AI Silicon

Ericsson GmbH • Austin (TX), Northern (KY)

Hybrid
USD 140,000 - 210,000
Senior DL Compiler Engineer — Fast Inference, Equity
Senior DL Compiler Engineer — Fast Inference, Equity

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 152,000 - 242,000