LLM & Generative AI Engineer: RAG & Fine-Tuning

Genixbit

Greater London

On-site

GBP 70,000 - 100,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Genixbit in London is seeking an AI Engineer to fine-tune open-source models for domain tasks, improve deployment pipelines, and develop context management and semantic search solutions.

The role requires hands-on experience with PyTorch, Hugging Face Transformers, and parameter-efficient fine-tuning, plus deploying models with vLLM, Ollama, or Triton. Strong software engineering practices are essential.

Qualifications

  • 3+ years of experience in NLP and Generative AI.
  • Hands-on PyTorch, Hugging Face Transformers, and PEFT/LoRA experience.
  • Experience deploying models with vLLM, Ollama, or Triton.

Responsibilities

  • Fine-tune open-source models (Llama, Mistral, Qwen) for domain functions.
  • Optimize deployment pipelines for low latency and high throughput.
  • Build advanced context management and semantic search solutions.
  • Implement prompt evaluation frameworks and guardrail architectures.

Skills

NLP / Generative AI
PyTorch
Hugging Face Transformers
PEFT / LoRA
Model deployment

Tools

vLLM
Ollama
Triton Inference Server

Job description

Genixbit in London is seeking an AI Engineer to fine-tune open-source models for domain tasks, improve deployment pipelines, and develop context management and semantic search solutions.

The role requires hands-on experience with PyTorch, Hugging Face Transformers, and parameter-efficient fine-tuning, plus deploying models with vLLM, Ollama, or Triton. Strong software engineering practices are essential.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

LLM & Generative AI Engineer: Fine-Tuning & RAG
LLM & Generative AI Engineer: Fine-Tuning & RAG

GenixBit Labs Pvt. Ltd. • Greater London

On-site
GBP 70,000 - 90,000
LLM & Generative AI Engineer
LLM & Generative AI Engineer

GenixBit Labs Pvt. Ltd. • Greater London

On-site
GBP 70,000 - 90,000
LLM & Generative AI Engineer
LLM & Generative AI Engineer

Genixbit • Greater London

On-site
GBP 70,000 - 100,000
Production ML Engineer - LLMs & Generative AI
Production ML Engineer - LLMs & Generative AI

Understanding Recruitment • Greater London

On-site
GBP 90,000 - 130,000
Equity
7% pension
Private healthcare
+1
Hybrid GenAI Journalist Engineer — LLM Fine-Tuning & RAG
Hybrid GenAI Journalist Engineer — LLM Fine-Tuning & RAG

Enfint • Greater London

Hybrid
GBP 70,000 - 120,000
Hybrid work pattern
Work From Anywhere up to 25 days/year
Private health insurance
+5
GenAI Data Scientist: LLM & RAG Architect (Hybrid, London)
GenAI Data Scientist: LLM & RAG Architect (Hybrid, London)

TechYard • Greater London

Hybrid
GBP 90,000 - 140,000
Senior AI Engineer: LLMs & Finetuning for Finance
Senior AI Engineer: LLMs & Finetuning for Finance

9fin • Greater London

Hybrid
GBP 65,000 - 90,000
Competitive Salary
Equity
Pension
+5
GenAI Data Scientist - LLM & RAG Expert (Hybrid)
GenAI Data Scientist - LLM & RAG Expert (Hybrid)

Experis • Greater London

Hybrid
GBP 81,000 - 99,000
Life assurance
Employee share purchase programme
Flexible hybrid working
LLM Researcher: Architect & Optimizer (London, Hybrid)
LLM Researcher: Architect & Optimizer (London, Hybrid)

OpenAI • Greater London

Hybrid
GBP 80,000 - 110,000
Relocation assistance
Hybrid work model
Hybrid LLM Engineer — AI/ML, RAG & Vision
Hybrid LLM Engineer — AI/ML, RAG & Vision

Ultralytics • Greater London

Hybrid
GBP 85,000 - 110,000
Equity
Hybrid schedule
24 days vacation
+2