ml engineer for LLM products

Kiefer

United States

On-site

USD 140,000 - 210,000

Full time

3 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Remote work
Relocation support
Athens office access

Job summary

KIEFER ищет инженера ML для разработки Sophea AI: обучение LLM с нуля, продвинутые ML-пайплайны и оптимизация моделей на GPU. Вакансия удалённая с возможной релокацией в Афины и офис в Афинах. Требуется сильный опыт в LLM, Python, PyTorch и продакшн-практиках.

Компания предоставляет доступ к передовым AI-инфраструктурам, участие в конференциях и развитие в рамках амбициозного греческого AI-экосистемы. Ожидаются навыки квантования и оптимизации инференса, работа с мультиимплементациями и

Qualifications

  • Опыт работы с LLMs, включая дообучение и оценку.
  • ML-инжиниринг: Python, PyTorch, Docker, продакшн-практики.
  • Опыт развёртывания моделей, оптимизации вывода, квантования и нагрузки на GPU.
  • Умение строить production-grade ML-системы, не только прототипы.
  • Носитель польского языка на уровне носителя.
  • Желателен: pre-training, ASR, мультиязычные модели, MLOps и вклад в open-source.

Responsibilities

  • Разрабатывать и улучшать Sophea AI, включая обучение LLM с нуля.
  • Строить production-grade ML-пайплайны для инференса, развёртывания и мониторинга.
  • Оптимизировать производительность моделей: задержку, пропускную способность и затраты.
  • Работать с наборами данных, экспериментами и метриками для качества и域ной производительности.

Skills

LLMs
Python
PyTorch
Production ML
Model serving
Inference optimization
Quantization
GPU workloads
LM fine-tuning

Tools

vLLM
SGLang
NVIDIA Triton
TensorRT
TGI

Job description

Описание:

KIEFER is building Greece's integrated AI ecosystem, spanning renewable energy infrastructure, AI systems, robotics, and enterprise applications. Founded in 2014, it has delivered 600MW+ of energy projects and is developing sovereign AI infrastructure, enterprise AI products, and physical AI systems for Greece and Southeast Europe.

Задачи:
  • Develop and continuously improve Sophea AI, including LLM training from scratch, fine-tuning, evaluation, and model improvement
  • Build production-grade ML pipelines for inference, serving, deployment, monitoring, and model lifecycle management
  • Optimize production model performance, including latency, throughput, cost efficiency, quantization, and GPU workload usage
  • Work with datasets, experiments, benchmarks, and evaluation methods to improve language model quality and domain-specific performance
Требования:
  • Strong hands-on experience with LLMs, including fine-tuning, evaluation, and performance improvement
  • Strong ML engineering background, including Python, PyTorch, Docker, and production ML practices
  • Experience with model serving, inference optimization, quantization, GPU workloads, and frameworks such as vLLM, SGLang, NVIDIA Triton, TensorRT, TGI, or similar tools
  • Ability to build production-grade ML systems, not only research prototypes, scripts, basic RAG applications, or high-level AI integrations
  • Native-level Polish proficiency
  • Будет плюсом: pre-training and training experience, ASR systems, speech models or speech-to-text pipelines, non-English or multilingual language models or low-resource language adaptation, MLOps infrastructure, experiment tracking, model serving pipelines, GPU workload management, contributions to open-source ML projects or published AI/ML research
Условия:
  • Competitive compensation package aligned with talent benchmarks
  • Remote work option; relocation support is available for candidates open to working from the Athens office
  • Hands-on work on Sophea AI, described as one of the most ambitious Greek-focused AI products in the market
  • Access to AI-related conferences, certifications, internal knowledge sharing, and advanced AI infrastructure through Kiefer's strategic collaboration with NVIDIA Engineering-first culture, high autonomy, low bureaucracy, and space to build meaningful AI products
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior ML Engineer (LLM)
Senior ML Engineer (LLM)

Kiefer • Kentucky

Remote
USD 120,000 - 180,000
Remote work option
Relocation support
NVIDIA ecosystem access
ML Team Lead (Sophea AI)
ML Team Lead (Sophea AI)

Kiefer • Kentucky

Remote
USD 150,000 - 230,000
Remote work option
Relocation support to Athens
NVIDIA ecosystem access
Remote ML Engineer for Production-Grade LLMs
Remote ML Engineer for Production-Grade LLMs

Kiefer • United States

Remote
USD 140,000 - 210,000
Remote work
Relocation support
Athens office access
Senior LLM Engineer - Production ML Systems (Remote)
Senior LLM Engineer - Production ML Systems (Remote)

Kiefer • Kentucky

Remote
USD 120,000 - 180,000
Remote work option
Relocation support
NVIDIA ecosystem access
Senior Software Engineer (Backend focus)
Senior Software Engineer (Backend focus)

Kiefer • Kentucky

Remote
USD 78,000 - 123,000
Paid annual leave (28 weekdays)
Private health insurance
Conference budget
ml engineer for LLM inference
ml engineer for LLM inference

Cast AI • United States

On-site
USD 180,000 - 270,000
Equity 10%
Annual hackathon
Learning budget
Remote Lead ML Engineer: LLMs, ASR & Production Systems
Remote Lead ML Engineer: LLMs, ASR & Production Systems

Kiefer • Kentucky

Remote
USD 150,000 - 230,000
Remote work option
Relocation support to Athens
NVIDIA ecosystem access
ai engineer for LLM applications
ai engineer for LLM applications

EPAM • United States

Remote
USD 140,000 - 200,000
ml engineer for real-time entertainment
ml engineer for real-time entertainment

Mayflower • United States

Remote
USD 140,000 - 190,000
ml engineer for llm performance optimization
ml engineer for llm performance optimization

NDA • United States

On-site
USD 150,000 - 230,000
Leave + wellness days
Training coverage
Health stipend
+1