Research Engineer: Inference Architect for Fast LLMs Remote

Kog

Paris

Hybride

EUR 60 000 - 90 000

Plein temps

14 jours+

Recevez plus de réponses des employeurs

Envoyez un CV adapté au poste en quelques minutes.

Avantages offerts par ce poste

Direct access to AMD and NVIDIA GPUs
Creative and decision-influencing environment
Remote-friendly working model

Résumé du poste

Kog is seeking an innovative architect to design and optimize model architectures focused on inference behavior. You'll experiment with cutting-edge AI technologies in a dynamic team setting.

In this role, you'll influence our model development processes while conducting research that enhances execution speed. A collaborative environment allows your technical judgment to shape key decisions with direct impact on system evolution.

Qualifications

  • Experience adapting or modifying existing model architectures.
  • Understanding of communication structure and layer dependencies affecting inference behavior.
  • Fluency in Transformers and MoE with depth to reason across trade-offs.

Responsabilités

  • Design new model architecture variants with execution constraints as input.
  • Explore inference‑aware architectural variants and discover scalable compounds.
  • Own the post-training pipeline for open-weight models optimized for inference speed.
  • Scale large MoE models and analyze communication patterns at inference.
  • Write and present findings in research papers at top venues.

Connaissances

Model design
Architectural optimization
Post-training methods
AI problem-solving

Formation

PhD or equivalent in a relevant field

Description du poste

Kog is seeking an innovative architect to design and optimize model architectures focused on inference behavior. You'll experiment with cutting-edge AI technologies in a dynamic team setting.

In this role, you'll influence our model development processes while conducting research that enhances execution speed. A collaborative environment allows your technical judgment to shape key decisions with direct impact on system evolution.

Obtenez votre examen gratuit et confidentiel de votre CV.
ou faites glisser et déposez votre fichier ici.
Similar jobs

Postes similaires à comparer

Research Engineer
Research Engineer

Kog • Paris

Hybride
EUR 60 000 - 90 000
Direct access to AMD and NVIDIA GPUs
Creative and decision-influencing environment
Remote-friendly working model
LLM Architecture Research Engineer
LLM Architecture Research Engineer

Ensimag Alumni • Paris

Sur place
EUR 50 000 - 80 000
Research Engineer (LLM Architecture)
Research Engineer (LLM Architecture)

Ensimag Alumni • Paris

Sur place
EUR 50 000 - 80 000
Research Engineer - FAIR, SGT
Research Engineer - FAIR, SGT

Meta • Paris

Sur place
EUR 90 000 - 130 000
AI Engineer - Self-Hosted LLMs & High-Performance Inference
AI Engineer - Self-Hosted LLMs & High-Performance Inference

In Tandem • Paris

Sur place
EUR 60 000 - 80 000
ML Inference Engineer | Accelerate AI on Custom Accelerator
ML Inference Engineer | Accelerate AI on Custom Accelerator

Arago • Paris

Sur place
EUR 90 000 - 130 000
Stock options
Healthcare coverage
Pension contributions
+2
Research Engineer — Multimodal AI & Agentic Systems
Research Engineer — Multimodal AI & Agentic Systems

H Company • Paris

Hybride
EUR 70 000 - 100 000
Competitive salary
Opportunities for professional growth
Collaborative and dynamic work environment
GPU Engineer
GPU Engineer

Ensimag Alumni • Paris

Sur place
EUR 50 000 - 70 000
Research Scientist: Core Learning & Reasoning for LLMs
Research Scientist: Core Learning & Reasoning for LLMs

Meta • Paris

Sur place
EUR 90 000 - 150 000
Staff ML Engineer: LLM/RAG Systems (Remote, Equity)
Staff ML Engineer: LLM/RAG Systems (Remote, Equity)

Valoris Fusion • Any-Martin-Rieux

Sur place
EUR 113 000 - 152 000
Remote-first
Equity
Performance bonus
+2