Senior ML Solutions Architect - Token Factory

Jobgether

España

A distancia

EUR 90.000 - 130.000

Jornada completa

Hace 5 días
Sé de los primeros/as/es en solicitar esta vacante
Generador de candidaturas

Una candidatura hecha para este puesto de trabajo: un currículum y una carta de presentación adaptados que responden directamente a la oferta.

Supera los filtros ATS

Ventajas ofrecidas por este puesto de trabajo

Competitive pay
Fully remote across Europe
Career growth & learning

Descripción de la vacante

Jobgether is hiring for an AI infra team focused on serverless platforms for running open-source LLMs in production. You will optimize inference workflows, fine-tune models, and design end-to-end AI architectures across multimodal models while collaborating with product and engineering teams.

This is a fully remote role across Europe with high ownership and impact. You will collaborate closely with customers to translate technical challenges into robust AI solutions, and you will help drive

Formación

  • 5+ years of professional experience with ML/AI systems, incl. 2+ years on LLMs.
  • Deep understanding of modern LLM ecosystems, architectures, and inference approaches.
  • Hands-on production experience deploying LLM workloads at scale.
  • Strong Python programming and production-oriented AI solutions.
  • Experience with prompt engineering, RAG, evaluation frameworks, and tooling.
  • Familiarity with cloud AI platforms (AWS, Google, Azure) and DevOps stacks.

Responsabilidades

  • Optimize LLM inference workflows across modalities to deliver business value.
  • Assist in supervised and RL-based fine-tuning to improve model quality.
  • Design LLM-powered solutions using serverless inference and open models.
  • Build production apps via LLM APIs across text, vision, audio, and domains.
  • Provide guidance on prompt engineering, RAG, deployment strategies.
  • Guide customers from PoC to production focusing on performance and reliability.
  • Collaborate with product and engineering to address platform gaps and roadmap.
  • Advise on model selection and tuning strategies based on use cases.

Conocimientos

5+ years ML/AI systems
2+ years LLMs & generative AI
Python programming
DevOps familiarity
Cloud AI platforms knowledge
Open-source ML tooling

Herramientas

vLLM
SGLang
TensorRT-LLM
Transformers
Docker
Kubernetes
Git
OpenAI/Anthropic APIs

Descripción del empleo

Join a fast-growing AI infrastructure team building a serverless platform for running and customizing open-source LLMs in production. Help customers move from AI prototypes to scalable, reliable production applications without building their own complex inference stacks. Design optimized inference workflows and customized LLM solutions across multiple models and modalities. Work hands-on with inference, fine-tuning, evaluation, prompt engineering, and retrieval-augmented generation. Partner directly with customers to understand technical challenges and translate them into effective AI architectures. Collaborate closely with product and engineering teams to turn customer feedback into platform improvements. Work remotely across Europe in an international environment focused on high-impact AI projects, technical ownership, and continuous innovation.

Accountabilities
  • Optimize LLM inference workflows across different modalities to deliver measurable business value and meet customer requirements.
  • Support customers with supervised and reinforcement-learning-based fine-tuning approaches to improve model quality and performance.
  • Design and implement LLM-powered solutions using serverless inference services and served open-source models.
  • Build production-ready applications using LLM APIs, including multimodal models covering text, vision, audio, and domain-specific use cases.
  • Provide technical guidance on prompt engineering, RAG architectures, model selection, inference optimization, and deployment strategies.
  • Guide customers through the transition from proof of concept to production, with a focus on performance, reliability, scalability, and cost efficiency.
  • Work closely with product and engineering teams to communicate customer needs, identify platform gaps, and contribute to roadmap development.
  • Help customers select appropriate models, inference configurations, and fine-tuning strategies based on their use cases and technical constraints.
  • Contribute to improving the platform and its capabilities by sharing practical insights from customer implementations and production workloads.
Requirements:
  • 5+ years of professional experience working with ML/AI systems, including at least 2 years focused specifically on LLMs and generative AI.
  • Deep understanding of the modern LLM ecosystem, including model architectures, inference approaches, and fine-tuning techniques.
  • Hands-on experience running LLMs in production, including deploying and operating inference workloads at scale.
  • Strong practical experience with LLM fine-tuning, including supervised fine-tuning, SFT, LoRA, and data preparation or curation; experience with reinforcement-learning-based fine-tuning is a strong advantage.
  • Experience building LLM evaluation frameworks, including task-specific benchmarks, offline and online evaluation pipelines, and LLM-as-a-judge approaches.
  • Practical experience with modern inference frameworks and ML libraries such as vLLM, SGLang, TensorRT-LLM, or Transformers.
  • Experience deploying LLM-powered applications through APIs from providers such as OpenAI or Anthropic, as well as open-source models.
  • Strong Python programming skills and the ability to develop practical, production-oriented AI solutions.
  • Excellent communication skills, with the ability to explain complex technical concepts clearly to customers, engineers, product teams, and other audiences.
  • Experience working with multimodal AI models, such as vision-language or speech models, is a plus.
  • Familiarity with DevOps technologies including Docker, Kubernetes, and Git is beneficial.
  • Contributions to open-source ML or AI projects are an additional advantage.
  • Familiarity with cloud AI platforms such as AWS SageMaker or Bedrock, Google Vertex AI, or Azure ML is welcome.
Benefits:
  • Competitive compensation.
  • Career growth and ongoing learning opportunities.
  • Flexible working arrangements with a high degree of autonomy and ownership.
  • Fully remote work opportunity from Europe.
  • Collaborative and innovative international working environment.
  • Opportunity to work on impactful AI infrastructure and production-grade LLM projects.
  • Exposure to advanced technologies across LLM inference, fine-tuning, evaluation, retrieval, and multimodal AI.
  • Opportunity to work closely with talented AI, engineering, product, and customer-facing teams.
  • Meaningful technical ownership and the opportunity to influence the evolution of an emerging AI platform.
  • Inclusive workplace committed to equal employment opportunities and a diverse working environment.
  • Reasonable accommodations are available during the application process where required.
  • Candidates must be authorized to work in the country where they apply and may need to provide proof of employment eligibility.
Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Artificial Intelligence Engineer
Artificial Intelligence Engineer

Migx • Barcelona

Híbrido
EUR 90.000 - 130.000
Hybrid work model
25 holiday days per year
Career development opportunities
+1
AI/ML Team Lead – Generative AI (LLMs, AWS)
AI/ML Team Lead – Generative AI (LLMs, AWS)

Provectus • España

A distancia
EUR 90.000 - 130.000
Signup bonus of 8,000 USD
Long-term B2B collaboration
Fully remote setup
+3
Ai Engineer
Ai Engineer

Capitole • Barcelona

Presencial
EUR 55.000 - 75.000
Hybrid work model (2 days onsite per週)
Private health insurance
€1200 per year training budget
+3
Senior ML Engineer (AI Research)
Senior ML Engineer (AI Research)

Jobgether • España

Híbrido
EUR 85.000 - 130.000
Competitive compensation
Learning opportunities
Flexible work
+1
AI Engineer
AI Engineer

Clarity • Madrid

Híbrido
EUR 70.000 - 105.000
Competitive pay
Location flexibility
Generous time off
+4
Senior Applied AI Solutions Engineer
Senior Applied AI Solutions Engineer

Jobgether • España

Presencial
EUR 174.000 - 304.000
Competitive USD salary
Autonomy and ownership
Global, collaborative team
+1
Backend Team Lead
Backend Team Lead

Acclaim AI • Barcelona

Presencial
EUR 90.000 - 130.000
Fully remote across Europe
Private English lessons via Preply
Company-paid subscriptions to top AI‑m
Lead Data & AI Engineer - GenAI & AI Platforms
Lead Data & AI Engineer - GenAI & AI Platforms

Value Crew • Barcelona

Híbrido
EUR 75.000 - 90.000
Hybrid work model
Private health insurance
Equity package
Applied Ai Engineer
Applied Ai Engineer

Jobgether • Málaga

Presencial
EUR 60.000 - 100.000
Healthcare coverage
Remote-first / Hybrid-friendly culture
Unlimited PTO
+1
Senior Machine Learning Engineer
Senior Machine Learning Engineer

Digital Waffle • España

Presencial
EUR 90.000 - 130.000