AI Framework & Inference Engineer

Black Sesame

San Jose, Northern (CA, KY)

Hybrid

USD 150,000 - 190,000

Full time

36 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Black Sesame seeks an experienced AI/ML engineer to deploy AI inference and end-to-end enablement, AI framework integration, model accuracy, and performance tuning. You will develop high-quality software that enables state-of-the-art AI inference on BST Intelligence processors.

Responsibilities include building deep learning infrastructure, data pipelines, and real-time deployment pipelines, with a focus on low-precision inference, quantization, and continual latency/accuracy improvements.

Qualifications

  • MS or Ph.D. in Computer Science, Computer Engineering, Applied Math, or related field.
  • Ability to work independently, define project goals and scope, and lead your own development effort.
  • Strong Python or C/C++ programming and software design skills, including debugging, performance analysis, and test design.
  • Experience with Deep Learning Frameworks such as PyTorch, TensorFlow, and MXNet.
  • Experience with model acceleration techniques such as deep learning quantization, model pruning, and model distillation.
  • Nice-to-have: Experience with numerical methods.
  • Nice-to-have: Knowledge of computer architecture.
  • Nice-to-have: Experience with AI compilers.
  • Nice-to-have: Experience in model deployment within the autonomous driving industry.

Responsibilities

  • Contribute to deep learning infrastructure, data pipelines, tools and workflows that lay the foundation for building AI at scale.
  • Writing software to deploy AI models and pipelines in real time applications (inference).
  • Apply low precision inference, quantization, and compression of DNNs.
  • Continuously improve inference latency, accuracy and power consumption of DNNs.
  • Stay up to date with the latest research and innovations in deep learning, implement and experiment with new ideas

Skills

Python
C/C++
Software design
PyTorch
TensorFlow
MXNet

Education

MS or Ph.D. in Computer Science / Computer Engineering / Applied Math or related field

Tools

PyTorch
TensorFlow
MXNet

Job description

Black Sesame seeks an experienced AI/ML engineer to deploy AI inference and end-to-end enablement, AI framework integration, model accuracy, and performance tuning. You will develop high-quality software that enables state-of-the-art AI inference on BST Intelligence processors.

Responsibilities include building deep learning infrastructure, data pipelines, and real-time deployment pipelines, with a focus on low-precision inference, quantization, and continual latency/accuracy improvements.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Compiler Engineer — Real-Time Inference & Optimization
AI Compiler Engineer — Real-Time Inference & Optimization

Black Sesame • San Jose (CA)

On-site
USD 120,000 - 170,000
Software Engineer, AI Framework
Software Engineer, AI Framework

Black Sesame • San Jose (CA), Northern (KY)

Hybrid
USD 150,000 - 190,000
High-Performance AI Inference Framework Engineer
High-Performance AI Inference Framework Engineer

ByteDance • San Jose (CA)

On-site
USD 150,000 - 230,000
Backend AI Engineer: Inference & Orchestration
Backend AI Engineer: Inference & Orchestration

ActAI • United States

On-site
USD 140,000 - 210,000
Software Engineer, AI Compiler
Software Engineer, AI Compiler

Black Sesame • San Jose (CA)

On-site
USD 120,000 - 170,000
Staff ML Engineer - Efficient Production Inference
Staff ML Engineer - Efficient Production Inference

ATBF Labs • San Francisco (CA)

Hybrid
USD 215,000 - 285,000
Equity
Health benefits
Senior Inference Engineer, AI Infrastructure & Production
Senior Inference Engineer, AI Infrastructure & Production

Hamilton Barnes • United States

On-site
USD 225,000 - 275,000
Full Benefits
Inference Performance Engineer: Accelerate LLMs & Cut Costs
Inference Performance Engineer: Accelerate LLMs & Cut Costs

US Health Partners, LLC • New York (NY), Northern (KY)

Hybrid
USD 150,000 - 210,000
Competitive compensation
Meaningful equity
US medical/dental/vision coverage for
+5
Software Engineer- Inference Performance
Software Engineer- Inference Performance

Baseten • United States

Remote
USD 150,000 - 210,000
Software Engineer, AI Inference - TS/SCI Eligible
Software Engineer, AI Inference - TS/SCI Eligible

Bitwise • Laurel (MD)

On-site
USD 198,000 - 233,000
Top salaries
PTO 3-5 weeks
Federal holidays paid
+5