Remote AI Kernel & Inference Engineer

Tether

Job

Híbrido

USD 180 000 - 260 000

Tempo integral

Há 12 dias
Gerador de candidaturas

Transforma esta função numa entrevista — um currículo e uma carta de apresentação criados à volta do que este empregador procura.

Ultrapassa os filtros ATS

Resumo da oferta

Tether is seeking a senior AI/ML engineer to advance model serving and inference for real-world fintech applications. You will optimize deployment pipelines across devices, from mobile to edge, focusing on low latency, memory efficiency, and scalable performance.

Ideal candidates hold a PhD in NLP/ML and have proven experience with diffusion models, vision transformers, and GPU-accelerated inference. Remote work is available with international collaboration and growth opportunities.

Qualificações

  • PhD in NLP or ML preferred and strong AI R&D background.
  • Deep expertise in designing and optimizing model serving pipelines.

Responsabilidades

  • Design and deploy state-of-the-art model serving architectures for high throughput and low latency across devices.
  • Build and monitor inference tests in simulated and live environments with KPIs.
  • Identify test datasets and simulations for real-world edge deployment challenges.
  • Analyze bottlenecks in serving pipelines and optimize memory/compute usage.
  • Collaborate with cross-functional teams to integrate serving frameworks into production pipelines.

Conhecimentos

Model Serving
Inference
Edge Devices
GPU Kernels
Metal Shading Language
Diffusion Models
Vision Transformers
Pruning
Quantization
Flash attention
KV Cache
Speculative Decoding

Formação académica

PhD in NLP

Ferramentas

Metal Shading Language (MSL)

Descrição da oferta de emprego

Tether is seeking a senior AI/ML engineer to advance model serving and inference for real-world fintech applications. You will optimize deployment pipelines across devices, from mobile to edge, focusing on low latency, memory efficiency, and scalable performance.

Ideal candidates hold a PhD in NLP/ML and have proven experience with diffusion models, vision transformers, and GPU-accelerated inference. Remote work is available with international collaboration and growth opportunities.

Obtém a tua avaliação gratuita e confidencial do currículo.

ou arrasta e larga o ficheiro aqui.

Similar jobs

Ofertas semelhantes que vale a pena comparar

AI Research Engineer (Kernel & Inference Optimization) - 100% Remote Worldwide
AI Research Engineer (Kernel & Inference Optimization) - 100% Remote Worldwide

Tether • Job

Híbrido
USD 180 000 - 260 000
On-Device AI Engineer for Real-Time Gaming
On-Device AI Engineer for Real-Time Gaming

Razer Inc. • França

Presencial
EUR 75 000 - 110 000
Senior AI Engineer
Senior AI Engineer

Razer Inc. • França

Presencial
EUR 75 000 - 110 000
AI/ML Engineer (Remote)
AI/ML Engineer (Remote)

Quik Hire Staffing • France

Presencial
EUR 55 000 - 90 000
Remote Biology Expert (PhD)
Remote Biology Expert (PhD)

Turing • França

Presencial
EUR 47 999 - 68 362
Senior AI Backend Engineer: Real-Time, Scalable Systems
Senior AI Backend Engineer: Real-Time, Scalable Systems

EfficientVision • Auvergne-Rhône-Alpes

Presencial
EUR 65 000 - 95 000
Remote Senior AI Engineer for Fintech ERP
Remote Senior AI Engineer for Fintech ERP

DualEntry • Eu

Presencial
EUR 123 033 - 175 762
Equity
Base salary
Remote-first team
+4
Research Engineer, Model Inference & Serving - Paris
Research Engineer, Model Inference & Serving - Paris

H Company • Paris

Presencial
EUR 60 000 - 90 000
Competitive salary
Opportunities for professional growth
Collaborative multicultural team
Senior Solutions Architect – Large Scale Neural Networks Inference
Senior Solutions Architect – Large Scale Neural Networks Inference

NVIDIA • Marseille

Presencial
EUR 120 000 - 190 000
Competitive salary
Benefits package
Remote AI/ML Engineer — Flexible Hours & Growth
Remote AI/ML Engineer — Flexible Hours & Growth

BairesDev • Paris

Teletrabalho
EUR 103 000 - 164 000
Remote work
USD salary
Hardware provided
+3