System-Level Reliability Researcher

Leuven

Heverlee

Sur place

EUR 90 000 - 120 000

Plein temps

14 jours+

Recevez plus de réponses des employeurs

Envoyez un CV adapté au poste en quelques minutes.

Résumé du poste

Imec’s Compute System Architecture (CSA) center of excellence in Leuven seeks an experienced System-Level Reliability Researcher to advance methodologies for evaluating and improving reliability of future compute architectures.

You will translate hardware fault and degradation observations into architectural fault models, run fault-injection and simulation, and quantify outcomes like Silent Data Corruption and availability loss, collaborating across CSA and with workload and circuit teams.

Qualifications

  • Developing system-level reliability assessment methods for processors, accelerators, memory subsystems, SoCs, and chiplet-based architectures.
  • Creating architectural fault models and abstractions based on reliability information provided by lower layers of the technology stack.
  • Building and using simulation, emulation, and fault-injection frameworks to study fault propagation, error masking, and application-level impact.
  • Quantifying reliability outcomes such as Silent Data Corruption, detected errors, service interruptions, availability loss, and lifetime-related degradation at the system level.
  • Evaluating Reliability, Availability, and Serviceability (RAS) mechanisms, including error detection, containment, recovery, redundancy, and graceful degradation techniques.
  • Studying workload-dependent reliability behavior under realistic execution conditions and operating profiles.
  • Connecting technology-informed fault characteristics with system architecture models to support reliability-aware design decisions.
  • Collaborating closely with colleagues within and outside CSA to interface with workload models, architectural simulators, circuit-level reliability data, and technology-level observations.
  • Contributing to research publications, partner discussions, etc.

Description du poste

Overview

Join imec’s center of excellence for hardware-software-technology co-design to define the future of system-level reliability in compute systems. Compute System Architecture (CSA) is a center of excellence at imec for hardware-software-technology co-design for future compute systems. We work in close collaboration with other expertise centers in imec specializing in applications, technology, circuits and design to innovate and pathfind next-generation compute system architectures across multiple domains – AI, HPC, Automotive, Space, and more. CSA has presence in 6 centers of IMEC, in Belgium, Netherlands, Germany, UK, USA and Qatar. This position is primarily for Leuven, Belgium.

What you will do

We are looking for an experienced System-Level Reliability Researcher to join our team. In this position, you will play a key role in developing methodologies to evaluate and improve the reliability of advanced compute architectures. You will focus on how hardware faults, degradation effects, and technology-level reliability observations translate into system-level behavior, application correctness, availability, and product lifetime.

This R&D role focuses on compute system architecture, reliability modeling, and resilience evaluation. You will collaborate with experts from the technology, device, circuit, and design domains to translate lower-level reliability information into architectural fault models, simulation inputs, and system-level reliability metrics.

Responsibilities
  • Developing system-level reliability assessment methods for processors, accelerators, memory subsystems, SoCs, and chiplet-based architectures.
  • Creating architectural fault models and abstractions based on reliability information provided by lower layers of the technology stack.
  • Building and using simulation, emulation, and fault-injection frameworks to study fault propagation, error masking, and application-level impact.
  • Quantifying reliability outcomes such as Silent Data Corruption, detected errors, service interruptions, availability loss, and lifetime-related degradation at the system level.
  • Evaluating Reliability, Availability, and Serviceability (RAS) mechanisms, including error detection, containment, recovery, redundancy, and graceful degradation techniques.
  • Studying workload-dependent reliability behavior under realistic execution conditions and operating profiles.
  • Connecting technology-informed fault characteristics with system architecture models to support reliability-aware design decisions.
  • Collaborating closely with colleagues within and outside CSA to interface with workload models, architectural simulators, circuit-level reliability data, and technology-level observations.
  • Contributing to research publications, partner discussions, etc.
Obtenez votre examen gratuit et confidentiel de votre CV.
ou faites glisser et déposez votre fichier ici.
Similar jobs

Postes similaires à comparer

System-Level Reliability Researcher
System-Level Reliability Researcher

Karlstad University • Heverlee

Sur place
EUR 90 000 - 130 000
System-Level Reliability Architect
System-Level Reliability Architect

Leuven • Heverlee

Sur place
EUR 90 000 - 120 000
System-Level Reliability Architect
System-Level Reliability Architect

Karlstad University • Heverlee

Sur place
EUR 90 000 - 130 000
Virtual Platform R&D Engineer
Virtual Platform R&D Engineer

Karlstad University • Heverlee

Sur place
EUR 75 000 - 110 000
Market salary + fringe benefits
Network Architect/Researcher
Network Architect/Researcher

Karlstad University • Heverlee

Sur place
EUR 90 000 - 120 000
Memory Sub-system R&D Engineer
Memory Sub-system R&D Engineer

Karlstad University • Heverlee

Sur place
EUR 70 000 - 110 000
Market-competitive salary
Fringe benefits
Network Architect/Researcher
Network Architect/Researcher

imec • Vlaams-Brabant

Sur place
EUR 90 000 - 130 000
Principal Architect - Compute Systems
Principal Architect - Compute Systems

Imec India Private Limited • Heverlee

Sur place
EUR 75 000 - 95 000
Market appropriate salary
Fringe benefits
Support for professional development
Failure Analysis Engineer
Failure Analysis Engineer

Imec India Private Limited • Heverlee

Sur place
EUR 65 000 - 85 000
Market appropriate salary
Fringe benefits
Investment in personal growth through imec.academy
R&D Group Manager – Compute Platforms & Silicon Prototypes
R&D Group Manager – Compute Platforms & Silicon Prototypes

Karlstad University • Heverlee

Sur place
EUR 110 000 - 140 000