Remote AI Research Engineer: Kernel & Inference Optimization

Lever, Inc.

México

A distancia

MXN 900.000 - 1.500.000

Jornada completa

Hace 2 días
Sé de los primeros/as/es en solicitar esta vacante
Generador de candidaturas

Transforma esta oferta en una entrevista — un currículum y una carta de presentación creados pensando en lo que quiere el empleador.

Supera los filtros ATS

Ventajas ofrecidas por este puesto de trabajo

Remote-first team
International collaboration
Cutting-edge AI research challenges

Descripción de la vacante

Lever, Inc. is seeking an AI Research Engineer (Kernel & Inference Optimization) based in Mexico.

You will work at the intersection of AI research, systems engineering, and high-performance model inference, focusing on latency, throughput, and memory efficiency across hardware environments. You will develop and optimize model-serving architectures, write custom GPU kernels (MSL), and collaborate with cross-functional teams to translate research into production edge applications.

Formación

  • Degree in Computer Science or related field; PhD in NLP/ML highly relevant.
  • Proven expertise in Metal Shading Language (MSL) and writing custom compute shaders.
  • Experience with low-level kernel optimization and inference optimization on mobile/edge devices.
  • Track record of measurable improvements in latency, throughput, and memory footprint.
  • Knowledge of modern model-serving architectures and optimization techniques for high-performance AI deployment.

Responsabilidades

  • Design and deploy advanced model-serving architectures optimized for high throughput and low latency.
  • Develop inference pipelines across diverse environments including edge devices.
  • Establish performance targets for latency, throughput, memory footprint, and reliability.
  • Build controlled inference benchmarks and track latency, throughput, and errors.
  • Create datasets and simulation scenarios to evaluate model performance.
  • Identify bottlenecks and apply system-level optimizations like batching and memory management.
  • Develop custom GPU kernels and compute shaders for mobile hardware (MSL).
  • Apply pruning, quantization, Flash Attention, KV caching, and speculative decoding.
  • Design distributed inference using tensor/pipeline/expert parallelism for large-scale GPUs.
  • Collaborate with cross-functional teams to integrate optimized inference frameworks.
  • Define evaluation methodologies and document experimental results.
  • Monitor production performance to identify improvement opportunities.

Conocimientos

MSL kernel programming
GPU kernel optimization
Inference optimization
Distributed inference
Benchmarking
English communication

Educación

PhD in NLP or Machine Learning

Herramientas

MSL (Metal Shading Language)

Descripción del empleo

Lever, Inc. is seeking an AI Research Engineer (Kernel & Inference Optimization) based in Mexico.

You will work at the intersection of AI research, systems engineering, and high-performance model inference, focusing on latency, throughput, and memory efficiency across hardware environments. You will develop and optimize model-serving architectures, write custom GPU kernels (MSL), and collaborate with cross-functional teams to translate research into production edge applications.

Consigue la evaluación confidencial y gratuita de tu currículum.

o arrastra y suelta tu archivo aquí

Similar jobs

Puestos de trabajo similares que vale la pena comparar

AI Research Engineer (Kernel & Inference Optimization)
AI Research Engineer (Kernel & Inference Optimization)

Lever, Inc. • México

A distancia
MXN 900.000 - 1.300.000
Remote-first team
International collaboration
Cutting-edge AI research challenges
Global MLOps Field Engineer — AI/ML Architect & Consultant
Global MLOps Field Engineer — AI/ML Architect & Consultant

Lever, Inc. • México

Presencial
MXN 1.592.000 - 2.300.000
Learning budget
Annual bonus
Distributed work
+1
AI Developer: LLM & RAG Systems Engineer
AI Developer: LLM & RAG Systems Engineer

Salvo Software LLC • México

Híbrido
MXN 900.000 - 1.300.000
Remote Senior AI Testing & Quality Engineer
Remote Senior AI Testing & Quality Engineer

Lever, Inc. • México

A distancia
MXN 2.602.000 - 4.196.000
VSOP equity
Fully remote work environment
Home office budget
+6
Enterprise AI Enablement Engineer
Enterprise AI Enablement Engineer

3050 Micron Semiconductor Mexico, S.de R.L. de C.V. • Jalisco

Presencial
MXN 400.000 - 600.000
Senior Research Scientist – Real-Time, Efficient AI Systems
Senior Research Scientist – Real-Time, Efficient AI Systems

adaption • Ciudad de México

Híbrido
MXN 1.200.000 - 2.400.000
Flexible work
Travel stipend
Lunch stipend
+1
AI & Data Scientist — Build LATAM's Advanced AI Engine
AI & Data Scientist — Build LATAM's Advanced AI Engine

Klar • Ciudad de México

Híbrido
MXN 600.000 - 900.000
Competitive salary
Klar stock options
15 days paid vacation
+5
AI Cloud Inference Market Development Intern
AI Cloud Inference Market Development Intern

QUALCOMM, Inc. • Ciudad de México

Presencial
MXN 67.000 - 112.000
Senior AI Engineer: LLMs, Autonomous Agents & AI Workflows
Senior AI Engineer: LLMs, Autonomous Agents & AI Workflows

Kms-Technology • Región Centro

Presencial
MXN 900.000 - 1.300.000
Mexican law benefits
Annual performance bonus
Major Medical Insurance
+5
AI/ML Engineer — Hybrid: Build Impactful AI Solutions
AI/ML Engineer — Hybrid: Build Impactful AI Solutions

FlexTal Staffing LLC • Región Centro

Híbrido
MXN 420.000 - 630.000
Christmas Bonus: 30 days
Major Medical Insurance: MXN 20,000,00
Dental Insurance
+9