Edge-Optimized AI Inference Kernel Engineer

Framework Ventures

United States

Remote

USD 180,000 - 260,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Job summary

Framework Ventures is seeking an experienced AI model serving engineer to advance inference pipelines across edge and cloud environments in the United States. You will design high‑throughput, low‑latency deployments for multi‑modal models and research novel serving strategies.

You will collaborate with cross‑functional teams to benchmark performance, optimize memory usage, and push the boundaries of real‑world AI capabilities on resource‑constrained devices.

Qualifications

  • Degree in Computer Science or related field; PhD in NLP/ML preferred.
  • Proven experience in low‑level kernel and inference optimization on mobile devices.
  • Strong understanding of modern model serving architectures and inference optimization techniques.
  • Experience writing GPU kernels for mobile devices and end‑to‑end inference pipelines.
  • Experience with empirical research and evaluation frameworks for optimization.

Responsibilities

  • Design and deploy state‑of‑the‑art model serving architectures with high throughput and low latency; optimize memory usage across diverse environments.
  • Build, run, and monitor controlled inference tests in simulated and live production environments; track latency, throughput, and memory metrics.
  • Identify and prepare high‑quality test datasets and scenarios for real‑world deployment challenges on low‑resource devices.
  • Analyze computational efficiency and diagnose bottlenecks in the serving pipeline; optimize for scalability and reliability on constrained systems.
  • Collaborate with cross‑functional teams to integrate optimized serving frameworks into edge/on‑device production pipelines; define success metrics.

Education

PhD in NLP / Machine Learning
BSc/ MSc in Computer Science

Tools

Metal Shading Language (MSL)
GPU kernels (mobile)
Inference optimization
Model serving architectures

Job description

Framework Ventures is seeking an experienced AI model serving engineer to advance inference pipelines across edge and cloud environments in the United States. You will design high‑throughput, low‑latency deployments for multi‑modal models and research novel serving strategies.

You will collaborate with cross‑functional teams to benchmark performance, optimize memory usage, and push the boundaries of real‑world AI capabilities on resource‑constrained devices.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote AI Research Engineer: Kernel & Inference
Remote AI Research Engineer: Kernel & Inference

Visa Hunt • Germany (OH)

On-site
USD 150,000 - 230,000
Senior Edge AI Inference Optimization Engineer
Senior Edge AI Inference Optimization Engineer

PVH (Tommy Hilfiger/Calvin Klein) • Santa Clara (CA)

On-site
USD 180,000 - 240,000
Intel Benefits
Remote AI Inference Engineer — Edge Model Deployment & Optimization
Remote AI Inference Engineer — Edge Model Deployment & Optimization

Quadric Inc. • Burlingame (CA)

On-site
USD 180,000 - 260,000
Health Care Plan (Medical, Dental & Vision)
Retirement Plan (401k, IRA)
Life Insurance (Basic, Voluntary & AD&D)
+7
Senior Edge Inference Optimization Engineer
Senior Edge Inference Optimization Engineer

Intel Corporation • Santa Clara (CA)

Hybrid
USD 195,000 - 361,000
Senior Edge Inference Optimization Engineer
Senior Edge Inference Optimization Engineer

Intel • Phoenix (AZ)

Hybrid
USD 195,000 - 362,000
Stock bonuses
Health insurance
Retirement plan
+1
Senior AI Kernel Engineer — Edge Inference & Optimization
Senior AI Kernel Engineer — Edge Inference & Optimization

quadric.io, Inc • Burlingame (CA)

On-site
USD 120,000 - 150,000
Health Care Plan (Medical, Dental & Vision)
Retirement Plan (401k, IRA)
Life Insurance (Basic, Voluntary & AD&D)
+7
Senior AI Inference Engineer — Edge ML, Porting Models
Senior AI Inference Engineer — Edge ML, Porting Models

quadric.io, Inc • Burlingame (CA)

On-site
USD 120,000 - 150,000
Health Care Plan (Medical, Dental & Vision)
Retirement Plan (401k, IRA)
Life Insurance (Basic, Voluntary & AD&D)
+7
Senior AI Inference Systems Engineer - Edge & Cloud
Senior AI Inference Systems Engineer - Edge & Cloud

Cloudflare • Austin (TX)

Hybrid
USD 180,000 - 230,000
Senior GPU AI Platform Engineer — Edge Inference (Equity)
Senior GPU AI Platform Engineer — Edge Inference (Equity)

NVIDIA AI • Seattle (WA)

On-site
USD 224,000 - 431,250
Equity
Benefits
Senior Edge AI Engineer - Remote
Senior Edge AI Engineer - Remote

Bright Vision Technologies • Santa Clara (CA)

On-site
USD 100,000 - 105,000
Remote work
Equal Opportunity Employer