PhD Position F/M Trustable Machine Learning : Analysis and Verification of Soft Automata

Inria

Rennes

Sur place

EUR 23 000 - 28 000

Plein temps

Il y a 8 jours
Générateur de candidature

Démarquez-vous pour ce poste — générez un CV et une lettre de motivation personnalisés en environ une minute.

Passez les filtres ATS

Avantages offerts par ce poste

Remboursement partiel des transports
Congés annuels + RTT
Télétravail possible après 6 mois
Équipement professionnel

Résumé du poste

Inria Rennes propose une thèse de doctorat F/M sur Trustable Machine Learning : Analyse et vérification des automates souples. Le sujet mêle apprentissage et méthodes formelles pour garantir les propriétés des modèles issus des données.

Le doctorant participera au projet SAIF et explorera des architectures telles que RNN, transformers et modèles d’espace d’états, au sein d’une équipe dynamique et ouverte à la collaboration internationale.

Qualifications

  • Forte formation en mathématiques et en méthodes formelles requise.
  • Expérience avec des bibliothèques et outils ML souhaitée.
  • Capacité à travailler sur des projets de recherche expérimentale et théorique.

Responsabilités

  • Risque théorique et expérimental sur les automates souples.
  • Conception et prototypage d’algorithmes ML et de vérification.
  • Rédaction de publications et participation à des conférences.

Connaissances

Mathématiques
Méthodes formelles
Bibliothèques ML

Formation

Master en CS (ou équivalent)

Outils

Python
PyTorch

Description du poste

PhD Position F/M Trustable Machine Learning : Analysis and Verification of Soft Automata

Fonction : Doctorant

The Inria center at the University of Rennes is one of eight Inria centers and has more than thirty research teams. The Inria center is a major and recognized player in the field of digital sciences. It is at the heart of a rich ecosystem of R&D and innovation, including highly innovative SMEs, large industrial groups, competitiveness clusters, research and higher education institutions, centers of excellence, and technological research institutes.

Location and environment:

The PhD will take place at INRIA Rennes (Brittany, France).

The candidate will be part of the collaborative project SAIF, “Safe AI through Formal methods,” (https://project.inria.fr/saif/), that involves renowned research labs in Computet Science : Inria, CEA-List, LIX, LaBRI, LMF, ENS Paris, ENS Saclay.

Salary includes health insurance and participation to public transportation expenses.

Topic:

Learning automata from their traces has long been addressed from a purely logical perspective (e.g. Angluin’s L* algorithm), until neural architectures offered an amazing alternative : ground breaking performances, summoning models at the boundary between the continuous world and the discrete world, leveraging probabilistic approaches... but providing no guarantees on the models produced by the learning algorithms !The objective of this thesis is to shed light on the properties of these “soft automata,” based on neural networks, by crossing perspectives from system theory, statistics, optimization and formal methods in order to provide guarantees on these dynamic systems, to understand their expressivity, their robustness to noise and attacks, and their sensitivity to data quality. The thesis will examine different architectures, from plain recurrent neural networks to gating and attention mechanisms, and up to more recent architectures like state space models or Mamba. The design of new neural architectures with better properties, and the design of jailbreaking and poisoning attacks to these models are also in the scope. More details below.

The adaptation of verification techniques to neural networks (NN) has (successfully) focused on a rather narrow topic : how robust is the output of a NN to perturbations on the input. Standard approaches are borrowed to static analysis, and perform reasonings at the scale of individul neurons. Besides scalability issues, these methods are oriented to classifiers and hardly adapt to models of dynamic systems. Mostly, they put aside the huge engineering effort that led to high performance neural architectures. This is the angle adopted here : exploiting this architecture to tailor verification approaches. Numerous neural architectures have been designed to identify dynamic systems from their traces. We focus here on the learning of automata from part of their language. These models are trained as predictors of the future, from positive examples only, and not as classifiers (deciding if some imput word is in the language or not). This makes them generative models, that could be used as surrogate of automata, whence the generic name of “soft automata” as these models compute with real numbers.

Recurrent neural networks (RNN) are the most natural neural architecture that comes to mind when one wants to learn an automaton. While trained with gradient descent, these objects have been shown to converge to discrete behaviors : their state space tends to form clusters which structure and properties are still under investigation. Similar behaviors appear with variants like LSTM or GRU, that inrtroduce gating mechanisms in order to prevent the fast memory decay of plain RNN. These emerging properties suggest that understanding the structuration of the state space of these models is key to address questions like their robustness to noise, to data quality and to attacks.

Independently, the success of transformers in text modeling/generation has motivated their adaptation to the larger domain of time series analysis. It is yet unclear if foundation models could emerge in that field, but successful attempts have been reported with rather simple architectures. The simplest is probably PatchTST, which abilities to learn automata remain to be explored (taking words in the language as time series). A possible research direction could be to identify how the attention mechanism and the sketching of patches in a time series combine to identify features in a sequence, and further to structure the state space of these models. Still with the aim of assessing their generalization abilities and their robustness to noise or attacks.

More recently, other architectures have been introduced under the generic term of “state space models,” like HiPPO or S4, and further Mamba. While originally addressing two limitations of transformers, a finite window context and a quadratic computational cost in the size of this window, they take inspiration from well known linear models in systems theory, and open the way to a more interpretable state space. A possible direction of the thesis could therefore be to explore the relevance of these models as surrogate automata, and again make use of their internal structure to design analysis and verification techniques.

The 3 research directions mentioned above will not all be explored at the same level. The topic will be adapted to the candidate. The ideal candidate should have a solid background in mathematics, a taste for formal methods and abilities for experimental work using standard machine learning libraries.

Requirements:

The ideal candidate should have a solid background in mathematics, a taste for formal methods and abilities for experimental work using standard machine learning libraries.

  • Gail Weiss, Yoav Goldberg, Eran Yahav : “On the Practical Computational Power of Finite Precision RNNs for Language Recognition,” 2018.
  • J. Michalenko, A. Shah, A. Verma, R. Baraniuk, S. Chaudhuri, A. Patel : “Representing Formal Languages : A Comparison Between Finite Automata and Recurrent Neural Networks,” ICLR 2019.
  • Zeyuan Allen-Zhu, Yuanzhi Li, “Physics of Language Models : Part 1, Learning Hierarchical Language Structures,” 2023, ICML 2024 tutorial.
  • Albert Gu, Tri Dao, “Mamba : Linear-Time Sequence Modeling with Selective State Spaces,” 2024, https://doi.org/10.48550/arXiv.2312.0075
  • Yuqi Nie, Nam H. Nguyen, Phanwadee Sinthong, Jayant Kalagnanam : “A time series is worth 64words : long-term forecasting with transformers,” ICLR 2023.

Main activities : the usual with PhD preparation

  • theoretical research
  • experimental research (prototyping original algorithms, use of machine learning libraries, experimental design, analysis of simulation results)
  • research paper writing (submission to journals and conferences),participation to conferences (includes traveling abroad)
  • participation to team meetings anf project meetings, oral presentation of results
  • thesis writing and thesis defense

Additional activities :

  • scientific training (a total of >100 hours is mandatory along the 3 years of the thesis)

Technical skills and level required : a Master in CS (or equivalent) is mandatory ; strong background in mathematics and theoretical computer science ; autonomy in software production (use of standard machine learning libraries) ; taste for formal methods

Languages : English, possibly French

Relational skills :ability to engage in informal personal or scientific exchanges and to establish connections with other students in the lab ; ability to speak to an audience (scientific presentation) ; scientific integrity ; reliability in work relations (conformance to work plan, regularity of work, commitment,...)

Other values appreciated : scientific creativity, strong curiosity

Avantages
  • Partial reimbursement of public transport costs
  • Leave: 7 weeks of annual leave + 10 extra days off due to RTT (statutory reduction in working hours) + possibility of exceptional leave (sick children, moving home, etc.)
  • Possibility of teleworking (after 6 months of employment) and flexible organization of working hours
  • Professional equipment available (videoconferencing, loan of computer equipment, etc.)
  • Social, cultural and sports events and activities
Obtenez votre examen gratuit et confidentiel de votre CV.
ou faites glisser et déposez votre fichier ici.
Similar jobs

Postes similaires à comparer

PhD Position F/M Pretrained models of multimodal neuroimaging for predicting individual cognition
PhD Position F/M Pretrained models of multimodal neuroimaging for predicting individual cognition

Inria • Palaiseau

Hybride
EUR 18 000 - 30 000
Remboursement des frais de transport
7 semaines de congés + RTT
Télétravail possible et organisation d
+2
Research engineer in AI security/fairness for specifying and implementing regulatory compliance tests for the EU AI Act
Research engineer in AI security/fairness for specifying and implementing regulatory compliance tests for the EU AI Act

Inria • Villeurbanne

Hybride
EUR 27 000 - 33 000
Partial transport reimbursement
Teleworking (90 days/yr)
Professional training
+3
PhD Position F/M Spatio-temporal analysis of remote sensing data at large scales
PhD Position F/M Spatio-temporal analysis of remote sensing data at large scales

Inria • Valbonne

Sur place
EUR 23 000 - 28 000
Remboursement partiel des frais de bus
Congés annuels 7 semaines + RTT
Télétravail possible
+4
PhD Position F/M Frugal Distributed Training with Volatile Resources
PhD Position F/M Frugal Distributed Training with Volatile Resources

Inria • Valbonne

Sur place
Partial reimbursement of public transport costs
7 weeks of annual leave + 10 extra days off
Possibility of teleworking
+3
Research Master Internship: New Algorithms for Automated Room Acoustic Diagnosis
Research Master Internship: New Algorithms for Automated Room Acoustic Diagnosis

Inria • Strasbourg

Hybride
EUR 6 700 - 10 000
Partial transport reimbursement
6 months teleworking possibility
Flexible working hours
+1
Postdoc in formal methods for control systems
Postdoc in formal methods for control systems

Enac Isae-Supaero • Toulouse

Sur place
EUR 40 000 - 50 000
Research Engineer - Automatic differentiation and control
Research Engineer - Automatic differentiation and control

Inria • France

Hybride
EUR 27 000 - 33 000
Public transport reimbursement
Extended leave (RTT) and annual leave
Teleworking and flexible hours
+4
Doctorant F/H Vers des modèles de diffusion efficaces
Doctorant F/H Vers des modèles de diffusion efficaces

Inria • Lyon

Hybride
EUR 20 000 - 27 000
Restauration subventionnée
Transports publics remboursés
Mutuelle et prévoyance
+1
PhD Position F/M Visualization of the Plausibility and Bias for Data Resources used in a Geographic Digital Twin
PhD Position F/M Visualization of the Plausibility and Bias for Data Resources used in a Geographic Digital Twin

Inria • Gif-sur-Yvette

Hybride
EUR 20 000 - 27 000
Partial transport reimbursement
7 weeks leave + RTT
Teleworking possible
+2
Contractual lecturer-researcher AI for Industry
Contractual lecturer-researcher AI for Industry

EURAXESS Ireland • Compiègne

Sur place