LLM & Generative AI Engineer

GenixBit Labs Pvt. Ltd.

Greater London

On-site

GBP 70,000 - 90,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

GenixBit Labs Pvt. Ltd. is seeking an AI Engineer in London specializing in LLM fine-tuning and Generative AI. You will work on optimizing and tailoring models for specific applications, ensuring high performance in low latency environments.

The ideal candidate has over 3 years of experience in Natural Language Processing, strong knowledge of PyTorch and Hugging Face Transformers, and a solid background in software engineering practices. Join us to innovate and drive AI solutions in diverse domains.

Qualifications

  • Minimum 3 years of experience in Natural Language Processing and Generative AI.
  • Experience with PyTorch and Hugging Face Transformers.
  • Strong software engineering practices and clean code.

Responsibilities

  • Fine-tune open-source models like Llama, Mistral, Qwen.
  • Optimize model deployment pipelines for efficiency.
  • Build context management and semantic search solutions.

Skills

Natural Language Processing
Generative AI
PyTorch
Hugging Face Transformers
Software Engineering

Tools

vLLM
Ollama
Triton Inference Server

Job description

About the Role

Join our AI Engineering division in London to specialize in LLM fine-tuning, retrieval-augmented generation (RAG), and hosting private models. You will be responsible for tailoring deep learning models to specialized domain tasks.

Key Responsibilities
  • Fine-tune open-source models (Llama, Mistral, Qwen) for specific domain functions
  • Optimize model deployment pipelines for low latency and high throughput
  • Build advanced context management and semantic search solutions
  • Implement prompt evaluation frameworks and guardrail architectures
Requirements
  • 3+ years of experience focusing on Natural Language Processing and Generative AI
  • Hands-on experience with PyTorch, Hugging Face Transformers, and parameter-efficient fine-tuning (PEFT/LoRA)
  • Experience deploying models with vLLM, Ollama, or Triton Inference Server
  • Strong background in software engineering best practices and clean code
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

LLM & Generative AI Engineer
LLM & Generative AI Engineer

Genixbit • Greater London

On-site
GBP 70,000 - 100,000
LLM & Generative AI Engineer: RAG & Fine-Tuning
LLM & Generative AI Engineer: RAG & Fine-Tuning

Genixbit • Greater London

On-site
GBP 70,000 - 100,000
LLM & Generative AI Engineer: Fine-Tuning & RAG
LLM & Generative AI Engineer: Fine-Tuning & RAG

GenixBit Labs Pvt. Ltd. • Greater London

On-site
GBP 70,000 - 90,000
AI Platform Engineer – LLM Infrastructure
AI Platform Engineer – LLM Infrastructure

Talenzon group • Greater London

On-site
GBP 60,000 - 80,000
Production ML Engineer - LLMs & Generative AI
Production ML Engineer - LLMs & Generative AI

Understanding Recruitment • Greater London

On-site
GBP 90,000 - 130,000
Equity
7% pension
Private healthcare
+1
Machine Learning Engineer | Competitive Compensation | LLMs
Machine Learning Engineer | Competitive Compensation | LLMs

Understanding Recruitment • Greater London

On-site
GBP 90,000 - 130,000
Equity
7% pension
Private healthcare
+1
Senior AI Engineer
Senior AI Engineer

Harnham • Greater London

Hybrid
GBP 90,000 - 140,000
Machine Learning Engineer
Machine Learning Engineer

VirtueTech Recruitment Group • Greater London

Hybrid
GBP 90,000 - 110,000
Hybrid work model
London office access
Generative AI Architect – Enterprise AI Systems
Generative AI Architect – Enterprise AI Systems

Talenzon group • Greater London

On-site
GBP 120,000 - 180,000
Researcher, Training - London
Researcher, Training - London

United States Digital Space LLC • Greater London

Hybrid
GBP 70,000 - 90,000
Relocation support
Hybrid work schedule