16 papers · ranked by Valyu relevance
В. И. Щербаков
Непрерывно возрастающие требования к скорости обработки информации приводит к необходимости иметь огромные вычислительные ресурсы. Современные суперкомпьютеры, представляющие собой массив элементарных компьютеров, несмотря на распараллеливание алгоритмов, имеют большие временные потери на пересылки информации между…
Junwei Zhou, HaoYun Xiao, Jianwen Xi, Qiuzhen Lin
—Distributed Arithmetic Coding (DAC) has emerged as a feasible solution to the Slepian-Wolf problem, particularly in scenarios with non-stationary sources and for data sequences with lengths ranging from small to medium. Due to the inherent decoding ambiguity in DAC, the number of candidate paths grows exponentially…
Lorenzo De Stefani
We present COPSIM a parallel implementation of standard integer multiplication for the distributed memory setting, and COPK a parallel implementation of Karatsuba's fast integer multiplication algorithm for a distributed memory setting. When using P processors, each equipped with a local memory, to compute the product…
James Hanlon, Simon Hollis, David C. May
—The ability to express a program as a hierarchical composition of parts is an essential tool in managing the complexity of software and a key abstraction this provides is to separate the representation of data from the computation. Many current parallel programming models use a shared memory model to provide data…
Marcin Pikus, Wen Xu, Kramer Gerhard
—A distribution matcher (DM) encodes a binary input data sequence into a sequence of symbols with a desired target probability distribution. Several DMs, including shell mapping and constant-composition distribution matcher (CCDM), have been successfully employed for signal shaping, e.g., in opticalfiber or 5G. The…
Rohan Yadav, Alex Aiken, Fredrik Kjølstad
We introduce DISTAL, a compiler for dense tensor algebra that targets modern distributed and heterogeneous systems. DISTAL lets users independently describe how tensors and computation map onto target machines through separate format and scheduling languages. The combination of choices for data and computation…
Vipul Gupta, Shusen Wang, Thomas A. Courtade, Kannan Ramchandran
—We propose OverSketch, an approximate algorithm for distributed matrix multiplication in serverless computing. OverSketch leverages ideas from matrix sketching and high-performance computing to enable cost-efficient multiplication that is resilient to faults and straggling nodes pervasive in low-cost serverless…
Victor Volfson
The paper considers the properties of pseudo stationarity in a broad sense and pseudo strong mixing for sequences of random variables corresponding to arithmetic functions. Assertions on this topic have been proven. The implementation of these properties for known arithmetic functions has been verified. The article…
Cosmin E. Oancea, Stephen M. Watt
We report on GPU implementations of block-level addition, subtraction, multiplication and division for midsize integers, with operands of $2^{15}$ to $2^{19}$ bits using the high-level functional language Futhark. Comparing with hand-written C++/CUDA versions and CGBN, we identify which functional constructs compile…
Cayo Dória
The purpose this article is to try to understand the mysterious coincidence between the asymptotic behavior of the volumes of the Moduli Space of closed hyperbolic surfaces of genus g with respect to the Weil-Petersson metric and the asymptotic behavior of the number of arithmetic closed hyperbolic surfaces of genus g.…
Benjamin Brock, Robert Cohn, Suyash Bakshi, Tuomas Kärnä + 5 more
'Jeongnim Kim' 'Mateusz Nowak' 'Łukasz Ślusarczyk' 'Kacper Stefanski' 'Timothy G. Mattson'] Data structures and algorithms are essential building blocks for programs, and distributed data structures, which automatically partition data across multiple memory locales, are essential to writing high-level parallel…
Łukasz Świerczewski
—Paper describes the theoretical and practical aspects of the proposed model that uses distributed computing to a global network of Internet communication. Distributed computing are widely used in modern solutions such as research, where the requirement is very high processing power, which can not be placed in one…
Kenneth Odoh
I am grateful to the numerous reading groups in Vancouver that spurred my interest in Distributed Systems. Despite my humble beginnings, I am now privileged to have developed into a seasoned Software Engineer. This book represents my opportunity to contribute back to society. Writing this book has been the most…
Eric B. Olsen
Residue Number Systems (RNS) offer efficient modular arithmetic and natural parallelism, but direct integer division in RNS remains a difficult and comparatively underdeveloped operation. This paper builds on the type-II division algorithm of Szabo and Tanaka and reformulates it for more efficient hardware…
Patrick Finnerty, Yoshiki Kawanishi, Tomio Kamada, Chikara Ohta
In this article we present our relocatable distributed collections library. Building on top of the AGPAS for Java library, we provide a number of useful intra-node parallel patterns as well as the features necessary to support the distributed nature of the computation through clearly identified methods. In particular…
B.R. Mehta, Jonti Talukdar, Sachin Gajjar
— Increasing development in embedded systems, VLSI and processor design have given rise to increased demands from the system in terms of power, speed, area, throughput etc. Most of the sophisticated embedded system applications consist of processors; which now need an arithmetic unit with the ability to execute complex…