LLM & Generative AI Engineer

Genixbit

Greater London

On-site

GBP 70,000 - 100,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Genixbit in London is seeking an AI Engineer to fine-tune open-source models for domain tasks, improve deployment pipelines, and develop context management and semantic search solutions.

The role requires hands-on experience with PyTorch, Hugging Face Transformers, and parameter-efficient fine-tuning, plus deploying models with vLLM, Ollama, or Triton. Strong software engineering practices are essential.

Qualifications

  • 3+ years of experience in NLP and Generative AI.
  • Hands-on PyTorch, Hugging Face Transformers, and PEFT/LoRA experience.
  • Experience deploying models with vLLM, Ollama, or Triton.

Responsibilities

  • Fine-tune open-source models (Llama, Mistral, Qwen) for domain functions.
  • Optimize deployment pipelines for low latency and high throughput.
  • Build advanced context management and semantic search solutions.
  • Implement prompt evaluation frameworks and guardrail architectures.

Skills

NLP / Generative AI
PyTorch
Hugging Face Transformers
PEFT / LoRA
Model deployment

Tools

vLLM
Ollama
Triton Inference Server

Job description

About the Role

Join our AI Engineering division in London to specialize in LLM fine-tuning, retrieval-augmented generation (RAG), and hosting private models. You will be responsible for tailoring deep learning models to specialized domain tasks.

Key Responsibilities
  • Fine-tune open-source models (Llama, Mistral, Qwen) for specific domain functions
  • Optimize model deployment pipelines for low latency and high throughput
  • Build advanced context management and semantic search solutions
  • Implement prompt evaluation frameworks and guardrail architectures
Requirements
  • 3+ years of experience focusing on Natural Language Processing and Generative AI
  • Hands-on experience with PyTorch, Hugging Face Transformers, and parameter-efficient fine-tuning (PEFT/LoRA)
  • Experience deploying models with vLLM, Ollama, or Triton Inference Server
  • Strong background in software engineering best practices and clean code
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

LLM & Generative AI Engineer
LLM & Generative AI Engineer

GenixBit Labs Pvt. Ltd. • Greater London

On-site
GBP 70,000 - 90,000
LLM & Generative AI Engineer: RAG & Fine-Tuning
LLM & Generative AI Engineer: RAG & Fine-Tuning

Genixbit • Greater London

On-site
GBP 70,000 - 100,000
LLM & Generative AI Engineer: Fine-Tuning & RAG
LLM & Generative AI Engineer: Fine-Tuning & RAG

GenixBit Labs Pvt. Ltd. • Greater London

On-site
GBP 70,000 - 90,000
AI Platform Engineer – LLM Infrastructure
AI Platform Engineer – LLM Infrastructure

Talenzon group • Greater London

On-site
GBP 60,000 - 80,000
Researcher, Training - London
Researcher, Training - London

United States Digital Space LLC • Greater London

Hybrid
GBP 70,000 - 90,000
Relocation support
Hybrid work schedule
GenAI Data Scientist: LLM & RAG Architect (Hybrid, London)
GenAI Data Scientist: LLM & RAG Architect (Hybrid, London)

TechYard • Greater London

Hybrid
GBP 90,000 - 140,000
Generative AI Architect – Enterprise AI Systems
Generative AI Architect – Enterprise AI Systems

Talenzon group • Greater London

On-site
GBP 120,000 - 180,000
Senior AI Engineer
Senior AI Engineer

0026 Checkout Technology Ltd • Greater London

On-site
GBP 85,000 - 110,000
Production AI Engineer: LLMs, NLP & Market Intelligence
Production AI Engineer: LLMs, NLP & Market Intelligence

Permutable.AI • Greater London

Hybrid
GBP 90,000 - 130,000
Generative AI Engineer: Build Production LLM Apps
Generative AI Engineer: Build Production LLM Apps

Harrington Starr • Greater London

On-site
GBP 70,000 - 110,000