12 papers · ranked by Valyu relevance
Liangliang Chang, Joshua Mack, Benjamin L. Willis, Xing Chen + 3 more
'John Brunhaver' 'Ali Akoglu' 'Chaitali Chakrabarti'] Abstract—In this study, we introduce a methodology for automatically transforming user applications in the radar and communication domain written in C/C++ based on dynamic profiling to a parallel representation targeted for a heterogeneous SoC. We present our…
Rajendra Purohit, K. R. Chowdhary, Sunıl Dutt Purohıt
—Arrival of multicore systems has enforced a new scenario in computing, the parallel and distributed algorithms are fast replacing the older sequential algorithms, with many challenges of these techniques. The distributed algorithms provide distributed processing using distributed file systems and processing units…
Parinaz Barakhshan, Rudolf Eigenmann
We compare automatically and manually parallelized NAS Benchmarks in order to identify code sections that differ. We discuss opportunities for advancing automatic parallelizers. We find ten patterns that pose challenges for current parallelization technology. We also measure the potential impact of advanced techniques…
Karame Mohammadiporshokooh, Steven R. Brandt, R. Tohid, Hartmut Kaiser
'Hartmut Kaiser'] Abstract. C++ Executors simplify the development of parallel algorithms by abstracting concurrency management across hardware architectures. They are designed to facilitate portability and uniformity of user-facing interfaces; however, in some cases they may lead to performance inefficiencies due to…
Temitayo Adefemi
—Parallelization has become a cornerstone of modern computing, influencing everything from high-performance supercomputers to everyday mobile devices. This paper presents a comprehensive guide on the fundamentals of parallelization that every computer scientist should know, beginning with a historical perspective that…
Madhav P. Desai
We investigate the utility of augmenting a microprocessor with a single execution pipeline by adding a second copy of the execution pipeline in parallel with the existing one. The resulting dual-hardware-threaded microprocessor has two identical, independent, single-issue in-order execution pipelines (hardware threads)…
Urmila Shrawankar, Mayuri Joshi
— In multi-core systems, various factors like inter-process communication, dependency, resource sharing and scheduling, level of parallelism, synchronization, number of available cores etc. influence the extent of possible High Performance Computing parallelization. These parameters if not managed to the root level…
Jesper Larsson Träff
These lecture notes are designed to accompany an imaginary, virtual, undergraduate, one or two semester course on fundamentals of Parallel Computing as well as to serve as background and reference for graduate courses on High-Performance Computing, parallel algorithms and shared-memory multiprocessor programming. They…
Donald S. Ene, V.I.E Anireh
- Evaluating how well a whole system or set of subsystems performs is one of the primary objectives of performance testing. We can tell via performance assessment if the architecture implementation meets the design objectives. Performance evaluations of several parallel algorithms are compared in this study. Both…
Denis Los, Igor Petushkov
Cores Authors: ['Denis Los' 'Igor Petushkov'] Abstract—Nowadays, latency-critical, high-performance applications are parallelized even on power-constrained client systems to improve performance. However, an important scenario of fine-grained tasking on simultaneous multithreading CPU cores in such systems has not been…
Brian Homerding, Atmn Patel, Enrico Armenio Deiana, Yian Su + 5 more
'Zujun Tan' 'Ziyang Xu' 'Bhargav Reddy Godala' 'David I. August' 'Simone Campanoni'] A compiler's intermediate representation (IR) defines a program's execution plan by encoding its instructions and their relative order. Compiler optimizations aim to replace a given execution plan (which instructions to execute and…
Tal Kadosh, Niranjan Hasabnis, Prema Soundararajan, Vy A. Vo + 4 more
Compilation Authors: ['Tal Kadosh' 'Niranjan Hasabnis' 'Prema Soundararajan' 'Vy A. Vo' 'Mihai Capotă' 'Nesreen K. Ahmed' 'Yuval Pinter' 'Gal Oren'] Manual parallelization of code remains a significant challenge due to the complexities of modern software systems and the widespread adoption of multi-core architectures.…