AI Inference & HPC Engineer – Performance & APIs

Topaz Labs

Emeryville (CA)

On-site

USD 90,000 - 150,000

Full time

2 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Full medical/dental/vision coverage
15 days PTO
5 personal days + holidays
401k matching
Equity and profit sharing

Job summary

Topaz Labs is seeking a Software Engineer to support and optimize our AI Engine. You will report to the Head of AI Engine and enhance performance, stability, and feature availability across internal APIs and production products.

You will bridge research and production, preparing new models for deployment and optimizing GPU/CPU workloads. The role requires hands-on experience with C/C++, API design, image processing, and multithreading, along with at least 1 year in a related field.

Qualifications

  • Hands-on experience with performance optimization, e.g. concurrency, multithreading, memory, speed, benchmarking, reliability
  • Experience architecting APIs for internal development
  • Hands-on experience implementing image processing or computational photography algorithms
  • Expert knowledge of C/C++
  • At least 1+ years of professional working experience in a related field

Responsibilities

  • Increase app performance, stability, and availability of new features
  • Simplify and improve the API of the framework
  • Serve as the technical bridge between the Deep Learning research team and production products
  • Help prepare new and updated models for production and assist with GPU/CPU optimization
  • Collaborate with hardware partners (NVIDIA, AMD, Intel, Apple) to optimize inference on their hardware

Skills

C/C++
API design
Performance optimization
Multithreading
GPU optimization
Image processing
Benchmarking

Tools

OpenCV
ffmpeg
ONNX
TensorRT

Job description

Topaz Labs is seeking a Software Engineer to support and optimize our AI Engine. You will report to the Head of AI Engine and enhance performance, stability, and feature availability across internal APIs and production products.

You will bridge research and production, preparing new models for deployment and optimizing GPU/CPU workloads. The role requires hands-on experience with C/C++, API design, image processing, and multithreading, along with at least 1 year in a related field.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Inference & HPC Engineer: Performance & API
AI Inference & HPC Engineer: Performance & API

Topaz Labs • Dallas (TX)

On-site
USD 110,000 - 150,000
100% covered medical/dental/vision
15 days annual PTO
401k matching
+1
AI Performance Engineer – HPC, ARM & Distributed Inference
AI Performance Engineer – HPC, ARM & Distributed Inference

EngineersOfAI • Austin (TX)

On-site
USD 90,000 - 120,000
AI Inference Platform Engineer
AI Inference Platform Engineer

BaseTen • New York (NY), San Francisco (CA)

On-site
USD 140,000 - 210,000
Equity
Medical, dental and vision insurance (
Flexible PTO including Winter Break
+4
AI Inference Performance & Scale Engineer
AI Inference Performance & Scale Engineer

AMD • San Jose (CA)

On-site
USD 150,000 - 210,000
Benefits at a glance
Applied AI Inference Performance Engineer
Applied AI Inference Performance Engineer

CoreWeave • San Francisco (CA)

On-site
USD 188,000 - 275,000
Medical benefits
401(k) with match
Flexible PTO
+5
Performance Engineer — AI Inference Systems
Performance Engineer — AI Inference Systems

Anthropic • San Francisco (CA)

Hybrid
USD 350,000 - 850,000
Visa sponsorship
Flexible hybrid work policy
GenAI Performance Engineer: C++, Python, Rust Expert
GenAI Performance Engineer: C++, Python, Rust Expert

Obsidian • New York (NY)

On-site
USD 110,000 - 170,000
Software Engineer, AI Inference / HPC
Software Engineer, AI Inference / HPC

Topaz Labs • Emeryville (CA)

On-site
USD 90,000 - 150,000
Full medical/dental/vision coverage
15 days PTO
5 personal days + holidays
+2
AI Inference Performance Engineer
AI Inference Performance Engineer

Fathom • United States

Remote
USD 120,000 - 170,000
AI Performance Engineer — C++/Rust Systems Optimizer
AI Performance Engineer — C++/Rust Systems Optimizer

Mercor • New York (NY)

On-site
USD 120,000 - 180,000