Cross-Tier and Distributed Caching and Data Management in Massively Distributed Systems

Inria

Rennes

Sur place

EUR 26 000 - 35 000

Plein temps

Il y a 5 jours
Soyez parmi les premiers à postuler

Recevez plus de réponses des employeurs

Envoyez un CV adapté au poste en quelques minutes.

Avantages offerts par ce poste

Public transport reimbursement
Annual leave and RTT days
Teleworking after 6 months
Equipment provided
Social, cultural and sports events
Vocational training

Résumé du poste

Inria Centre at Rennes University invites applications for an engineer position to study, implement, and evaluate cross-tier and distributed caching strategies for hierarchical multi-tier storage systems. You will collaborate with a PhD student on this topic within the MAGELLAN team.

The position is funded for one year with potential extension to 24 months, and is hosted at the Inria Center at Rennes University.

Qualifications

  • Solid background in distributed systems.
  • Experience building systems and tools.
  • Proficient in Python and Java.
  • Excellent written and oral English.
  • Strong collaboration and networking skills.

Responsabilités

  • Study novel cross-tier and distributed caching strategies, alongside supporting data management techniques
  • Prototype key caching strategies and data management techniques
  • Run experiments and Evaluation of results
  • Reporting, disseminating and presenting results.
  • Participate in project meetings and discussions with other partners.

Connaissances

Distributed systems
Python
Java
English
Collaboration

Description du poste

Cross-Tier and Distributed Caching and Data Management in Massively Distributed Systems

The Inria Centre at Rennes University is one of Inria's nine centres and has more than thirty research teams. The Inria Centre is a major and recognized player in the field of digital sciences. It is at the heart of a rich R&D and innovation ecosystem: highly innovative PMEs, large industrial groups, competitiveness clusters, research and higher education players, laboratories of excellence, technological research institute, etc.

Financial and working environment.

This engineerposition will be in the context of IPCEI-CIS (Important Project of Common European Interest – Next Generation Cloud Infrastructure and Services) DXP (Data Exchange Platform) project involving Amadeus and three Inria research teams (COAST, CEDAR and MAGELLAN). This project aims to design and develop an open-source management solution for a federated and distributed data exchange platform (DXP), operating in an open, scalable, and massively distributed environment (cloud-edge continuum). The position will be recruited and hosted at the Inria Center at Rennes University; and the work will be carried out within the MAGELLAN team in collaboration with other partners.

The position is for one year, with the possibility of an extension to 24 months.

Context:

The ever-growing number of services and Internet of Things (IoT) devices has resulted in data being distributed across different locations (regions and countries) and different storage tires. Additionally, data exhibits different usage patterns, including cold data (written once and never read), stream data (produced once and consumed by many), and hot data (written once and consumed by many). Furthermore, these data types have different performance and dependability requirements (e.g., low latency for data streams).

To ensure the reliability and improve the performance of data-intensive applications, data are either replicated or erasure-coded and distributed across different storage tiers, while frequently accessed data are stored on high-speed devices close to end users (i.e., cached). While much work has investigated data caching, data placement strategies (i.e., deciding what to cache), data movement, cache partitioning, cache eviction [1–8], and cost-efficient data redundancy techniques in caching systems [9], few efforts have focused holistic caching and data management when caches are distributed across heterogeneous platforms (from Edge to Cloud), utilize storage devices with varying performance and cost characteristics, and simultaneously serve diverse workloads, including traditional data services, serverless workflows, and data streaming.

The goal of this engineer position is to study, implement, and evaluate novel cross-tier and distributed caching strategies, alongside supporting data management techniques, for hierarchical multi-tier storage systems. The engineer will work closely with a PhD student on this topic.

References:

[1] Asit Dan and Don Towsley. 1990. An Approximate Analysis of the LRU and FIFO Buffer Replacement Schemes. SIGMETRICS Perform. Eval. Rev. 18, 1 (apr 1990), 143–152. https://doi.org/10.1145/98460.98525

[2] Marek Chrobak and John Noga. 1999. LRU is better than FIFO. Algorithmica 23 (02 1999), 180–185. https://doi.org/10.1007/PL00009255

[3] Blankstein, Aaron, Siddhartha Sen, and Michael J. Freedman. “Hyperbolic caching: Flexible caching for web applications.” 2017 USENIX Annual Technical Conference (USENIX ATC 17). 2017.

[4] Cristian Ungureanu, Biplob Debnath, Stephen Rago, and Akshat Aranya. 2013. TBF: A memory-efficient replacement policy for flash- based caches. In 2013 IEEE 29th International Conference on Data Engineering (ICDE). 1117–1128. https://doi.org/10.1109/ICDE.2013.6544902

[5] Orcun Yildiz, Amelie Chi Zhou, Shadi Ibrahim. 2018. Improving the Effectiveness of Burst Buffers for Big Data Processing in HPC Systems with Eley. Future Generation Computer Systems, Volume 86, 2018, Pages 308-318, ISSN 0167-739X, https://doi.org/10.1016/j.future.2018.03.029 .

[6] G. Aupy, O. Beaumont and L. Eyraud-Dubois, "Sizing and Partitioning Strategies for Burst-Buffers to Reduce IO Contention," 2019 IEEE International Parallel and Distributed Processing Symposium (IPDPS), Rio de Janeiro, Brazil, 2019,

[7] ZHANG, Yazhuo, YANG, Juncheng, YUE, Yao, et al. {SIEVE} is simpler than {LRU}: an efficient {Turn-Key} eviction algorithm for web caches. In : 21st USENIX Symposium on Networked Systems Design and Implementation (NSDI 24). 2024. p. 1229-1246.

[8] Juncheng Yang, Ziming Mao, Yao Yue, and K. V. Rashmi. GL-Cache: Group-level learning for efficient and high-performance caching. FAST’23, pages 115–134, 2023.

[9] RASHMI, K. V., CHOWDHURY, Mosharaf, KOSAIAN, Jack, et al. {EC-Cache}:{Load-Balanced},{Low-Latency} cluster caching with online erasure coding. In : 12th USENIX Symposium on Operating Systems Design and Implementation (OSDI 16). 2016. p. 401-417.

  • Study novel cross-tier and distributed caching strategies, alongside supporting data management techniques
  • Prototype key caching strategies and data management techniques
  • Run experiments and Evaluation of results
  • Reporting, disseminating and presenting results.
  • Participate in project meetings and discussions with other partners.
  • A solid background in the area of distributed systems
  • Experience with building systems and tools
  • Software development skills: Python andJava
  • Working experience in the areas of data management, storage and cachingsystemsare advantageous
  • Good collaborative and networking skills
  • Excellent written and oral communication in English
Avantages
  • Partial reimbursement of public transport costs
  • Leave: 7 weeks of annual leave + 10 extra days off due to RTT (statutory reduction in working hours) + possibility of exceptional leave (sick children, moving home, etc.)
  • Possibility of teleworking (after 6 months of employment) and flexible organization of working hours
  • Professional equipment available (videoconferencing, loan of computer equipment, etc.)
  • Social, cultural and sports events and activities
  • Access to vocational training
  • Social security coverage

Starting from €2,695 gross per month, based on your experience

Obtenez votre examen gratuit et confidentiel de votre CV.
ou faites glisser et déposez votre fichier ici.
Similar jobs

Postes similaires à comparer

PhD Position F/M Modeling and Simulation of Exascale Storage Systems
PhD Position F/M Modeling and Simulation of Exascale Storage Systems

Sabiagrik • Rennes

Sur place
Subsidized meals
Partial reimbursement of public transport costs
Possibility of teleworking (90 days per year)
+2
Research engineer on estimating environmental impacts of large-scale distributed platforms
Research engineer on estimating environmental impacts of large-scale distributed platforms

Inria • Lyon

Hybride
Partial transport reimbursement
Leave: 7 weeks + RTT days + special/EX
Teleworking 90 days/year
+3
Confirmed Software Development Engineer M/F SLICES-FR
Confirmed Software Development Engineer M/F SLICES-FR

Inria • Grenoble

Hybride
Partial reimbursement of public transport costs
7 weeks of annual leave plus additional days
Possibility of teleworking
+2
Cross-Tier Distributed Caching Engineer (Remote after 6mo)
Cross-Tier Distributed Caching Engineer (Remote after 6mo)

Inria • Rennes

Sur place
EUR 26 000 - 35 000
Public transport reimbursement
Annual leave and RTT days
Teleworking after 6 months
+3
PhD Position F/M Frugal Distributed Training with Volatile Resources
PhD Position F/M Frugal Distributed Training with Volatile Resources

Inria • Valbonne

Sur place
Partial reimbursement of public transport costs
7 weeks of annual leave + 10 extra days off
Possibility of teleworking
+3
Design of metasurfaces for terahertz and infrared applications
Design of metasurfaces for terahertz and infrared applications

Inria • France

Sur place
EUR 27 000 - 33 000
Teleworking option
Leave: generous vacation and RTT days
Professional equipment available
+1
Post-Doctoral Research Visit F/M Advanced HPC frameworks for real-time simulation of periodic m[...]
Post-Doctoral Research Visit F/M Advanced HPC frameworks for real-time simulation of periodic m[...]

Inria • Rennes

Hybride
EUR 46 000 - 55 000
Partial reimbursement of public travel
Leave: RTT and annual leave
Teleworking after 6 months
Research Engineer in 5G Broadcast Experimental Platforms and Validation.
Research Engineer in 5G Broadcast Experimental Platforms and Validation.

Inria • Rennes

Sur place
Partial public transport reimbursement
7 weeks annual leave + RTT days
Teleworking after 6 months
+4
Research Engineer Position - Java Developer - Corese Library
Research Engineer Position - Java Developer - Corese Library

Inria • France

Sur place
EUR 2 000 - 55 000
Partial reimbursement of public transport costs
7 weeks of annual leave
Possibility of teleworking
+2
Robotic perception engineer
Robotic perception engineer

Inria • Nice

Hybride
Partial reimbursement of public travel
Leave: 7 weeks + RTT
Teleworking possibility
+3