14 papers · ranked by Valyu relevance
Marcin Cieślik, Cameron Mura
PaPy, which stands for parallel pipelines in Python, is a highly flexible framework that enables the construction of robust, scalable workflows for either generating or processing voluminous datasets. A workflow is created from user-written Python functions (nodes) connected by 'pipes' (edges) into a directed acyclic…
Jiasi Shen, Martin Rinard, Nikos Vasilakis
We present KumQuat, a system for automatically generating data parallel implementations of Unix shell commands and pipelines. The generated parallel versions split input streams, execute multiple instantiations of the original pipeline commands to process the splits in parallel, then combine the resulting parallel…
Houming Wu, Ling Chen, Wenjie Yu
Large Models Training Authors: ['Houming Wu' 'Ling Chen' 'Wenjie Yu'] With the increasing scale of models, the need for efficient distributed training has become increasingly urgent. Recently, many synchronous pipeline parallelism approaches have been proposed to improve training throughput. However, these approaches…
Nikos Vasilakis, Κωνσταντίνος Καλλάς, Konstantinos Mamouras, Achilles Benetopoulos + 1 more
'Achilles Benetopoulos' 'Lazar Cvetković'] This paper presents PaSh, a system for parallelizing POSIX shell scripts. Given a script, PaSh converts it to a dataflow graph, performs a series of semantics-preserving program transformations that expose parallelism, and then converts the dataflow graph back into a…
Shivam Handa, Κωνσταντίνος Καλλάς, Nikos Vasilakis, Martin Rinard
We present a dataflow model for modelling parallel Unix shell pipelines. To accurately capture the semantics of complex Unix pipelines, the dataflow model is order-aware, i.e., the order in which a node in the dataflow graph consumes inputs from different edges plays a central role in the semantics of the computation…
Camillo Lugaresi, Jiuqiang Tang, Hadon Nash, Chris McClanahan + 10 more
'Esha Uboweja' 'Michael L. Hays' 'Fan Zhang' 'Chuo-Ling Chang' 'Ming Guang Yong' 'Ju Hyun Lee' 'Wan-Teh Chang' 'Wei Hua' 'Manfred Georg' 'Matthias Grundmann'] Building applications that perceive the world around them is challenging. A developer needs to (a) select and develop corresponding machine learning algorithms…
Cheng-Hsiang Chiu, Tsung‐Wei Huang, Zizheng Guo, Yibo Lin
Pipeline is a fundamental parallel programming pattern. Mainstream pipeline programming frameworks count on data abstractions to perform pipeline scheduling. This design is convenient for data-centric pipeline applications but inefficient for algorithms that only exploit task parallelism in pipeline. As a result, we…
Aaron Harlap, Deepak Narayanan, Amar Phanishayee, Vivek Seshadri + 3 more
'Nikhil R. Devanur' 'Gregory R. Ganger' 'Phillip B. Gibbons'] PipeDream is a Deep Neural Network (DNN) training system for GPUs that parallelizes computation by pipelining execution across multiple machines. Its pipeline parallel computing model avoids the slowdowns faced by data-parallel training when large models…
Juan Cabral, B. Sánchez, M. Beroiz, M. Domínguez + 3 more
'S. Gurovich' 'Pablo M. Granitto'] Data processing pipelines represent an important slice of the astronomical software library that include chains of processes that transform raw data into valuable information via data reduction and analysis. In this work we present Corral, a Python framework for astronomical pipeline…
H. M. R. D. B. Nava, S. Radhakrishnan, Roshan Ragel
- Application Specific Instruction-set Processor (ASIP) is one of the popular processor design techniques for embedded systems which allows customizability in processor design without overly hindering design flexibility. Multi-pipeline ASIPs were proposed to improve the performance of such systems by compromising…
Daniel Ruprecht
The paper introduces an OpenMP implementation of pipelined Parareal and compares it to a standard MPI-based implementation. Both versions yield essentially identical runtimes, but, depending on the compiler, the OpenMP variant consumes about 7% less energy. However, its key advantage is a significantly smaller memory…
Vivekanandan Balasubramanian, Antons Treikalis, Ole Weidner, Shantenu Jha
'Shantenu Jha'] Abstract—There are many science applications that require scalable task-level parallelism, support for flexible execution and coupling of ensembles of simulations. Most high-performance system software and middleware, however, are designed to support the execution and optimization of single tasks.…
Yasset Pérez‐Riverol, Roberto Vera Alvarez
Parallel and distributed application design is a major area of interest in the domain of high performance scientific and industrial computing. Over the years, various approaches have been proposed to aid parallel program developers to modeling their applications. In this paper it will be used some concepts from agile…
Patrick Mukala
| Article Info | ABSTRACT | | --- | --- | | | A myriad of applications ranging from engineering and scientific | | | simulations, image and signal processing as well as high-sensitive data | | | retrieval demand high processing power reaching up to teraflops for their | | | efficient execution. While a standard serial…