Autonomous Driving AI Inference & Framework Engineer

Black Sesame Technologies Inc

San Jose (CA)

On-site

USD 140,000 - 210,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Black Sesame Technologies Inc in San Jose is seeking an experienced AI/ML engineer to deploy AI inference and end-to-end enablement on BST Intelligence processors. You will integrate AI frameworks and optimize model accuracy and latency for real-time applications.

Ideal candidates have MS/PhD, strong Python/C++, experience with PyTorch and TensorFlow, and knowledge of quantization, pruning and distillation to drive efficient deployment.

Qualifications

  • MS or PhD in CS, CE, Applied Math or related field.
  • Strong Python or C++ programming and software design.
  • Experience with PyTorch or TensorFlow.
  • Experience with model quantization, pruning, and distillation.
  • Ability to work independently and lead development efforts.

Responsibilities

  • Contribute to deep learning infrastructure, data pipelines, tools and workflows for AI at scale.
  • Write software to deploy AI models and pipelines in real-time applications (inference).
  • Apply low-precision inference, quantization, and compression of DNNs.
  • Continuously improve inference latency, accuracy and power consumption of DNNs.
  • Stay up to date with the latest research in deep learning and experiment with new ideas.

Skills

Python
C/C++ programming
Deep learning
Perf analysis
Software design

Education

MS or PhD in CS/CE/Applied Math

Tools

PyTorch
TensorFlow

Job description

Black Sesame Technologies Inc in San Jose is seeking an experienced AI/ML engineer to deploy AI inference and end-to-end enablement on BST Intelligence processors. You will integrate AI frameworks and optimize model accuracy and latency for real-time applications.

Ideal candidates have MS/PhD, strong Python/C++, experience with PyTorch and TensorFlow, and knowledge of quantization, pruning and distillation to drive efficient deployment.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Software Engineer, AI Framework
Software Engineer, AI Framework

Black Sesame Technologies Inc • San Jose (CA)

On-site
USD 140,000 - 210,000
ML Frameworks Engineer — Inference & Accelerator Systems
ML Frameworks Engineer — Inference & Accelerator Systems

Meta • Bellevue (WA)

On-site
USD 122,000 - 181,000
Graduate Backend Inference Engine Engineer
Graduate Backend Inference Engine Engineer

ByteDance • San Jose (CA)

On-site
USD 128,000 - 256,000
Medical insurance
Dental insurance
Vision insurance
+8
AI Compiler Engineer: ML Accelerator Optimizer
AI Compiler Engineer: ML Accelerator Optimizer

Black Sesame Technologies Inc • San Jose (CA)

On-site
USD 100,000 - 130,000
Senior ML Engineer: Quantized Inference & Pipelines
Senior ML Engineer: Quantized Inference & Pipelines

NVIDIA AI • Redmond (WA)

On-site
USD 140,000 - 200,000
Equity
Benefits
Senior Staff Engineer — AI Inference Co-Design Lead
Senior Staff Engineer — AI Inference Co-Design Lead

Samsung Semiconductor • San Jose (CA)

Hybrid
USD 189,000 - 301,000
Medical/Dental/Vision
401k
Wellness apps
+2
Senior AI Infrastructure Engineer, Inference & Optimization
Senior AI Infrastructure Engineer, Inference & Optimization

Didi Labs • San Jose (CA)

On-site
USD 170,000 - 351,000
Senior AI Inference & Optimization Infrastructure Engineer
Senior AI Inference & Optimization Infrastructure Engineer

DiDi • San Jose (CA)

On-site
USD 170,000 - 351,000
AI Infrastructure & ML Systems Intern
AI Infrastructure & ML Systems Intern

ByteDance • San Jose (CA)

On-site
USD 97,000 - 138,000
Health insurance
Housing allowance
Paid holidays
High-Performance AI Inference Framework Engineer
High-Performance AI Inference Framework Engineer

ByteDance • San Jose (CA)

On-site
USD 150,000 - 230,000