12 papers · ranked by Valyu relevance
Cristian Ramon-Cortes, Ramon Amela, Jorge Ejarque, Philippe Clauss + 1 more
'Rosa Badia'] Abstract—The last improvements in programming languages, programming models, and frameworks have focused on abstracting the users from many programming issues. Among others, recent programming frameworks include simpler syntax, automatic memory management and garbage collection, which simplifies code…
Alcides Fonseca, Bruno Cabral, João Rafael, Ivo Correia
This work proposes a new approach for achieving such goal. We created a new parallelizing compiler that analyses the read and write instructions, and control-flow modifications in programs to identify a set of dependencies between the instructions in the program. Afterwards, the compiler, based on the generated…
Tal Kadosh, Niranjan Hasabnis, Prema Soundararajan, Vy A. Vo + 4 more
Compilation Authors: ['Tal Kadosh' 'Niranjan Hasabnis' 'Prema Soundararajan' 'Vy A. Vo' 'Mihai Capotă' 'Nesreen K. Ahmed' 'Yuval Pinter' 'Gal Oren'] Manual parallelization of code remains a significant challenge due to the complexities of modern software systems and the widespread adoption of multi-core architectures.…
Krzysztof Stuglik, Piotr Listkiewicz, Mateusz Kulczyk, Marcin Pietroń
'Marcin Pietroń'] Manual translation of the algorithms from sequential version to its parallel counterpart is time consuming and can be done only with the specific knowledge of hardware accelerator architecture, parallel programming or programming environment. The automation of this process makes porting the code much…
Garip Kusoglu, Bérenger Bramas, Stéphane Genaud
Currently, multi/many-core CPUs are considered standard in most types of computers including, mobile phones, PCs or supercomputers. However, the parallelization of applications as well as refactoring/design of applications for efficient hardware usage remains restricted to experts who have advanced technical knowledge…
Clément Aubert, Thomas Rubiano, Neea Rusch, Thomas Seiller
This work explores an unexpected application of Implicit Computational Complexity (ICC) to parallelize loops in imperative programs. Thanks to a lightweight dependency analysis, our algorithm allows splitting a loop into multiple loops that can be run in parallel, resulting in gains in terms of execution time similar…
Ashwani Jha, K. M. Flurchick, Marwan Bikdash, Dukka B. KC
Internally symmetric proteins are proteins that have a symmetrical structure in their monomeric single-chain form. Around 10-15% of the protein domains can be regarded as having some sort of internal symmetry. In this regard, we previously published SymD (symmetry detection), an algorithm that determines whether a…
Alvaro Estebanez, Diego R. Llanos, David Orden, Belen Palop + 1 more
'Rafael Sachetto Oliveira'] Loops are a rich source of parallelism. Unfortunately, many loops cannot be safely parallelized at compile time because the compiler is not able to guarantee that there will be no dependence violations. Thread-Level Speculation (TLS) techniques, either hardware or software-based, allow the…
Muhammad Adnan, Faisal Aslam, Zubair Nawaz, Syed Mansoor Sarwar + 1 more
'Maciej Huk'] Nowadays, a typical processor may have multiple processing cores on a single chip. Furthermore, a special purpose processing unit called Graphic Processing Unit (GPU), originally designed for 2D/3D games, is now available for general purpose use in computers and mobile devices. However, the traditional…
Bahman Arasteh, Seyed Salar Sefati, Huseyin Kusetogullari, Farzad Kiani + 3 more
Efficient task scheduling remains a key challenge in High-Performance Computing and Internet of Things (IoT) systems, where the sequential execution of nested loops often limits parallelism. This paper proposes a hybrid approach that dynamically parallelizes nested loops in heterogeneous IoT environments. The suggested…
Clément Flint, Ludovic Paillat, Bérenger Bramas, Muhammad Aleem
High-performance computing (HPC) relies increasingly on heterogeneous hardware and especially on the combination of central and graphical processing units. The task-based method has demonstrated promising potential for parallelizing applications on such computing nodes. With this approach, the scheduling strategy…
Daniel Langenkämper, Tobias Jakobi, Dustin Feld, Lukas Jelonek + 2 more
'Alexander Goesmann' 'Tim W. Nattkemper'] Within the recent years clock rates of modern processors stagnated while the demand for computing power continued to grow. This applied particularly for the fields of life sciences and bioinformatics, where new technologies keep on creating rapidly growing piles of raw data…