Senior Data Scientist, Reinforcement Learning

Resaro AI

München

Vor Ort

EUR 90.000 - 120.000

Vollzeit

14 Tage+

Erhalte mehr Antworten von Arbeitgebern

Versende in nur wenigen Minuten einen passgenauen Lebenslauf.

Benefits dieser Stelle

Work on mission-critical AI systems
Collaborative expert team environment
Opportunity to shape product direction

Zusammenfassung

A leading AI technology company in Munich seeks a Senior Expert/VP Reinforcement Learning to architect their AI Test, Evaluation, Verification, and Validation (TEVV) product suite. You'll lead and mentor a global team to develop next-generation frameworks for reinforcement learning applications in Autonomous Driving and Robotics. Candidates should hold a Master’s or Ph.D. in a related field and have a strong background in RL algorithms and machine learning. This role provides an opportunity to shape AI testing standards in critical industries.

Qualifikationen

  • Proven track record in developing and implementing RL and ML algorithms.
  • Demonstrated understanding of the RL framework, including bandit and trust-region approaches.
  • Knowledge of AI/ML/RL lifecycle and testing limitations.

Aufgaben

  • Implement the RL validation prototype to expose agent vulnerabilities.
  • Lead a global team of AI researchers and engineers.
  • Define the vision and technical roadmap for RL TEVV.

Kenntnisse

Robot Reinforcement Learning
Machine Learning Algorithms
Bayesian Machine Learning
Requirements Gathering
Stakeholder Communication

Ausbildung

Master/Ph.D. in Robot Reinforcement Learning or related field

Jobbeschreibung

Resaro builds advanced AI testing software to help organizations verify, validate, and trust their most critical AI systems — from computer vision to generative AI and autonomous systems. Our mission is to ensure that AI technologies deployed in real-world, high-stakes environments are robust, explainable, and secure.

We work closely with our customers through embedded delivery teams who operate on-site or in close collaboration. These teams tailor solutions to specific mission needs, helping organizations — especially in the public safety and national security sectors — evaluate and improve the performance of their AI-enabled systems.

About the Role:

As the Senior Expert/ VP Reinforcement Learning, you will be the primary architect of our AI Test, Evaluation, Verification, and Validation (TEVV) product suite for reinforcement learning systems. You will lead the development of next-generation AI testing and assurance frameworks with applications in Autonomous Driving and Robotics. Your mission is to scale our capabilities in Reinforcement Learning, to ensure autonomous agents are safe, robust, and explainable in the field.

Key Responsibilities
  • Independently implement Resaro’s RL validation prototype to expose agent instability and vulnerability in a mission-critical and complex environment.
  • Scale, lead and mentor a global, cross-functional, high-performing team of AI researchers and engineers, drawing on experience steering organizations of 30+ experts.
  • Define the long-term vision and technical roadmap for RL TEVV, focusing on validating RL algorithms and learned policies in complex environments with mission-critical applications across system control, autonomous vehicles, and robotics.
  • Advance methods for learning probabilistic reward functions from human feedback (RLHF) to align AI behavior with mission goals.
  • Partner with Product Management to translate product vision, customer problems, and market opportunities into end‑to‑end solution architecture and technical roadmaps that support a product-led growth strategy.
Must-Have Skills and Experience
  • Master / Ph.D. in Robot Reinforcement Learning or a closely related field.
    • Proven track record in developing and implementing novel RL and ML algorithms, e.g. research or commercial implementation.
    • Demonstrated deep theoretical understanding of and practical experience with the RL framework, including bandit setting, (in-)finite horizon setting, on- and off-policy RL, and trust-region RL approaches.
  • Experience in Bayesian Machine Learning and probabilistic models.
  • Understanding of AI/ML/RL lifecycle and the state-of-the-art approaches and limitations of testing and validating complex use cases.
  • Strong skills in requirements gathering, stakeholder communication, and solution scoping.
Nice-to-Have
  • Experience with fully differentiable deep learning for highly unstable systems.
  • Experience with Active Learning and RLHF.
  • Background in model compression and pruning for deploying large RL models onto edge devices.
  • Hands-on experience with Bayesian Meta-Learning to reduce training time and absolute error in complex models.
  • A strong portfolio of innovation, including multiple successful paper submissions at conferences like NeurIPS, ICML, ICLR, IROS, ICRA, CoRL, and a deep patent history (e.g., 17+ patents).
  • Experience spearheading global AI initiatives and delivering AI solutions for both B2G (Unmanned Systems) and B2B (IoT) sectors.
  • Demonstrated success in leading cross-functional teams to deliver technical solutions.
  • Knowledge of deployment constraints in high-security or classified environments.
  • Prior exposure or experience with directly engaging senior stakeholders from Director to C-suite level.
  • Prior security clearance at Government CONFIDENTIAL and above.
Why Join Resaro
  • Work on mission-critical AI systems in defence, aerospace, and public safety.
  • Help define the future of AI testing and assurance in real-world environments.
  • Collaborate with a tight-knit, expert team working at the intersection of AI, systems engineering, and policy.
  • Shape product direction while being close to the operational reality of AI deployments.

Resaro is an Equal Opportunity Employer. We respect each individual and support the diverse cultures, perspectives, skills and experiences within our teams.

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.
oder ziehe deine Datei hierhin.
Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Senior AI Engineer - Reinforcement Learning
Senior AI Engineer - Reinforcement Learning

Resaro • München

Vor Ort
EUR 70.000 - 100.000
Opportunity to work on mission-critical AI systems
Collaborate with an expert team
Shape product direction
Senior/AI Engineer (m/f/d)
Senior/AI Engineer (m/f/d)

Resaro • München

Vor Ort
EUR 65.000 - 85.000
Solution Architect - Implementation Lead
Solution Architect - Implementation Lead

Resaro AI • München

Vor Ort
EUR 70.000 - 90.000
Collaborative work environment
Chance to work on mission-critical AI systems
Influence product direction
Solution Architect - Implementation Lead
Solution Architect - Implementation Lead

Resaro • München

Vor Ort
EUR 110.000 - 160.000
AI Quality Architect
AI Quality Architect

Resaro • München

Vor Ort
EUR 70.000 - 90.000
R&D Engineering Manager, AI Evaluation
R&D Engineering Manager, AI Evaluation

Resaro • München

Vor Ort
EUR 120.000 - 180.000
Senior LLM Scientist (m/f/d)
Senior LLM Scientist (m/f/d)

Resaro • München

Vor Ort
EUR 80.000 - 110.000
(Senior) AI / Reinforcement Learning Engineer (m/f/d)
(Senior) AI / Reinforcement Learning Engineer (m/f/d)

Agile Robots SE • München

Vor Ort
EUR 110.000 - 140.000
Health, mobility & learning benefits
Modern office in Munich
Competitive equity package
Senior Software Engineer
Senior Software Engineer

Resaro • München

Vor Ort
EUR 65.000 - 85.000
Senior Forward Deployed Engineer
Senior Forward Deployed Engineer

develop • Berlin

Hybrid
EUR 90.000 - 130.000