Edge AI Inference Engineer: Kernel & Serving

Tether

United Kingdom

Remote

GBP 80,000 - 100,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Tether is seeking a talented individual to join our AI model team. The role focuses on driving innovation in model serving and inference architectures for advanced AI systems. You will lead efforts to optimize model deployment and strategies for efficient performance across diverse real-world applications.

Key responsibilities include designing and deploying state-of-the-art serving architectures, managing inference tests, and analyzing and optimizing efficiency for resource-constrained environments. Strong educational background and expertise in modern AI methodologies are essential.

Qualifications

  • Strong knowledge of Metal Shading Language (MSL) and experience with custom compute shaders.
  • Proven experience in low‑level kernel optimizations and inference optimization on mobile devices.
  • Deep understanding of modern model serving architectures.

Responsibilities

  • Design and deploy state-of-the-art model serving architectures.
  • Build, run, and monitor controlled inference tests.
  • Analyze computational efficiency and optimize serving infrastructure.

Skills

Experience in NLP
Knowledge of Metal Shading Language (MSL)
Kernel optimizations
Inference optimization
Writing GPU kernels
Distributed inference systems

Education

PhD in NLP, Machine Learning
Degree in Computer Science

Job description

Tether is seeking a talented individual to join our AI model team. The role focuses on driving innovation in model serving and inference architectures for advanced AI systems. You will lead efforts to optimize model deployment and strategies for efficient performance across diverse real-world applications.

Key responsibilities include designing and deploying state-of-the-art serving architectures, managing inference tests, and analyzing and optimizing efficiency for resource-constrained environments. Strong educational background and expertise in modern AI methodologies are essential.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Edge AI Model Compression & Quantization Engineer
Edge AI Model Compression & Quantization Engineer

Tether.io • United Kingdom

On-site
GBP 60,000 - 90,000
AI Research Engineer — LLM & Multimodal Architectures
AI Research Engineer — LLM & Multimodal Architectures

Tether.io • United Kingdom

On-site
GBP 80,000 - 120,000
Inference Systems Performance Engineer for AI Serving
Inference Systems Performance Engineer for AI Serving

adaption • Greater London

On-site
GBP 90,000 - 130,000
Flexible work
Lunch stipend
Well-Being benefits
Senior AI Inference Platform Architect
Senior AI Inference Platform Architect

CoreWeave • Greater London

On-site
GBP 150,000 - 190,000
Family-level Medical Insurance
Family-level Dental Insurance
Generous Pension Contribution
+4
GenAI Infra Engineer: Real-Time GPU Serving & Fine-Tuning
GenAI Infra Engineer: Real-Time GPU Serving & Fine-Tuning

United States Digital Space LLC • Greater London

Hybrid
GBP 120,000 - 170,000
Inference Performance Engineer
Inference Performance Engineer

adaption • Greater London

On-site
GBP 90,000 - 130,000
Flexible work
Lunch stipend
Well-Being benefits
AI Inference Engineer — GPU-Optimized Rust/Python (Equity)
AI Inference Engineer — GPU-Optimized Rust/Python (Equity)

CVFine by Instrovate Technologies • Greater London

On-site
GBP 70,000 - 90,000
AI Research Engineer, Multimodal & LLM Architect
AI Research Engineer, Multimodal & LLM Architect

Tether Operations Limited • United Kingdom

On-site
GBP 70,000 - 110,000
Founding AI Inference Engineer – Scale & Serving Expert
Founding AI Inference Engineer – Scale & Serving Expert

Fuse Energy • Greater London

On-site
GBP 150,000 - 210,000
Equity sign-on bonus
Biannual bonus
Fully expensed tech
+1
Robotics AI Inference Engineer
Robotics AI Inference Engineer

Thehumanoid • Greater London

On-site
GBP 65,000 - 85,000
23 days annual leave
Fully funded private healthcare
Pension scheme with 8% contribution
+2