Applied Scientist (LLM)

SQUAD Ukraine Limited

Wrocław

On-site

PLN 334,800 - 558,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive salary packages
Guaranteed paid vacation
Private medical insurance

Job summary

A leading tech firm is seeking an experienced Applied Scientist to develop high-performance Generative AI features for Cloud and Edge environments. The role focuses on optimizing local inference, tuning models, and collaborating with product teams to integrate LLM features. Ideal candidates should have over 3 years of experience in Machine Learning, strong Python skills, and proficiency with PyTorch and Hugging Face. Benefits include competitive salaries, paid vacation, and opportunities for professional growth.

Qualifications

  • 3+ years of commercial experience in Machine Learning, NLP, or LLM.
  • Strong knowledge of Python3 and modern text-processing libraries.
  • Practical experience with Retrieval-Augmented Generation (RAG) or Fine-tuning.

Responsibilities

  • Drive transition from research to production through model optimization.
  • Design advanced methods in prompt orchestration and workflows.
  • Collaborate with Product and Data Engineering for LLM integration.

Skills

Machine Learning
NLP/LLM
Python3
PyTorch
Hugging Face
Data Analysis

Tools

NumPy
pandas
Docker
Kubernetes

Job description

Our distributed team is looking for an experienced Applied Scientist with a strong background in Large Language models to develop high-performance Generative AI features across Cloud and Edge environments.

Job Summary

In this role you will drive the transition from research to production by optimizing local inference through model compression and quantization for private, real-time Edge performance, while also engineering scalable RAG architectures and multi-agent systems for Cloud deployment. Your daily responsibilities encompass the full research lifecycle, including formulating hypotheses, generating synthetic datasets, fine-tuning LLMs, and validating safety and alignment, ultimately culminating in technical reports.

Responsibilities and Duties
  • Design and implement advanced methods in prompt orchestration, fine-tuning (SFT/RLHF/DPO), and autonomous agentic workflows
  • Curate high-quality training data from large-scale text and multi-modal sources
  • Identify patterns in model hallucinations and visualize evaluation metrics for clear interpretation
  • Tune hyperparameters and improve inference speed/accuracy through PEFT (LoRA/QLoRA) and advanced prompt engineering
  • Collaborate with Product and Data Engineering teams to seamlessly integrate LLM features into the broader ecosystem
  • Track and report progress using industry-standard benchmarks (MMLU, HumanEval, etc.) and custom internal KPIs
  • Stay at the forefront of the field (e.g., State Space Models, new Transformer variants) and evaluate cutting-edge techniques for production readiness
  • Engage in continuous technical growth and mentor junior colleagues to elevate the team's expertise
Qualifications and Skills
  • 3+ years of commercial experience in Machine Learning, with a specific focus on the NLP or LLM domain
  • Strong knowledge of Python3, NumPy, pandas, and modern text-processing libraries, PyTorch and Hugging Face (Transformers, PEFT, Accelerate)
  • Proficiency in PEFT/LoRA and Reinforcement Learning techniques
  • Deep understanding of attention mechanisms, tokenization, context window management, and embedding spaces
  • Practical experience in at least one of the following: Retrieval-Augmented Generation (RAG), Fine-tuning, or Agentic frameworks
  • Proven ability to manage and analyze massive datasets (>100GB) across text, image, and audio formats
  • Hands-on experience crafting high-fidelity datasets and building robust data pipelines
  • Expertise in prompt engineering, agentic framework design, and LLM pipeline orchestration
  • Experience deploying LLMs to production environments using Triton Inference Server, vLLM, TGI, or ONNX
  • Good written and spoken English
Nice to have
  • Practical experience with Pinecone, Weaviate, Milvus, or Chroma
  • Advanced quantization (GGUF, AWQ, EXL2), pruning, and knowledge distillation
  • Experience with LangChain, LlamaIndex, or AutoGen
  • Basic understanding of web/client-server architecture and streaming API responses (Asyncio, aiohttp)
  • Familiarity with RAGAS, DeepEval, or G-Eval
  • Experience using Docker, Kubernetes, and cloud GPU orchestration (e.g., Run:ai, Lambda Labs)
  • Knowledge of C++, Triton, or CUDA for custom kernel development
We offer multiple benefits that include
  • The environment of equal opportunities, transparent and value-based corporate culture, and an individual approach to each team member
  • Competitive salary packages with performance-based annual reviews
  • Employment via Contract of Employment (UoP) in complete alignment with Polish Labour Law
  • Guaranteed paid vacation, public holidays, and medical leaves as per statutory regulations
  • Continuous growth and development opportunities through internal knowledge hubs, corporate courses, and free English classes
  • Comprehensive private medical insurance to supplement your standard NFZ coverage.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Engineer (RAG & On Prem LLMs) RAG, LLM, Python, PyTorch, LangChain, Neo4j, Docker, Kubernete[...]
AI Engineer (RAG & On Prem LLMs) RAG, LLM, Python, PyTorch, LangChain, Neo4j, Docker, Kubernete[...]

Diverse CG Sp. z o.o. Sp.k. • Warszawa

On-site
PLN 212,000 - 340,000
AI Engineer (RAG & On Prem LLMs) RAG, LLM, Python, PyTorch, LangChain, Neo4j, Docker, Kubernete[...]
AI Engineer (RAG & On Prem LLMs) RAG, LLM, Python, PyTorch, LangChain, Neo4j, Docker, Kubernete[...]

DCG Poland • Warszawa

On-site
PLN 40,000 - 60,000
Senior AI/ML Engineer - Remote, MLOps & RAG
Senior AI/ML Engineer - Remote, MLOps & RAG

Formamind sp. z o.o • Warszawa

On-site
PLN 180,000 - 300,000
Architect ML/GenAI IRC296971
Architect ML/GenAI IRC296971

GlobalLogic • Kraków

On-site
PLN 180,000 - 340,000
Chief Forward Deployed Engineer
Chief Forward Deployed Engineer

EPAM Systems • Poland

Hybrid
PLN 260,000 - 380,000
Hybrid work model
Opportunity to work abroad up to 60,00
Relocation opportunities
+2
Formamind AI Engineer
Formamind AI Engineer

Formamind sp. z o.o • Warszawa

Remote
PLN 180,000 - 300,000
Flexible working hours
100% remote setup
Conference attendance
+3
Lead AI Engineer (Python)
Lead AI Engineer (Python)

NFQ Group • Kraków

On-site
Private medical care
Sports package
Life insurance
+2
Data Scientist
Data Scientist

MDPI USA • Kraków

On-site
PLN 80,000 - 120,000
Private medical care (fully covered)
MultiSport card (partially covered)
Team building activities
Lead AI Engineer
Lead AI Engineer

EPAM Systems • Poland

Hybrid
PLN 180,000 - 250,000
Hybrid work design
Relocation opportunities
English language classes
+2
Data Scientist
Data Scientist

MDPI • Kraków

On-site
PLN 50,000 - 80,000
Private medical care
MultiSport card
Team building activities