13 papers · ranked by Valyu relevance
David Castells-Rufas, Albert Saà-Garriga, Jordi Carrabina
The growing capacity of integration allows to instantiate hundreds of soft-core processors in a single FPGA to create a reconfigurable multiprocessing system. Lately, FPGAs have been proven to give a higher energy efficiency than alternative platforms like CPUs and GPGPUs for certain workloads and are increasingly used…
Lauritz Thamsen, Yehia Elkhatib, Paul Harvey, Syed Waqar Nabi + 2 more
Scientific research in many fields routinely requires the analysis of large datasets, and scientists often employ workflow systems to leverage clusters of computers for their data analysis. However, due to their size and scale, these workflow applications can have a considerable environmental footprint in terms of…
János Végh
This paper reinterprets Amdahl's law in terms of execution time and applies this simple model to supercomputing. The systematic discussion results in a quantitative measure of computational efficiency of supercomputers and supercomputing applications, explains why supercomputers have different efficiencies when using…
János Végh
Today we live in the age of artificial intelligence and machine learning; from small startups to HW or SW giants, everyone wants to build machine intelligence chips, applications. The task, however, is hard: not only because of the size of the problem: the technology one can utilize (and the paradigm it is based upon)…
Stephen J. Maher, Ted K. Ralphs, Yuji Shinano
Empirical studies are fundamental in assessing the effectiveness of implementations of branchand-bound algorithms. The complexity of such implementations makes empirical study difficult for a wide variety of reasons. Various attempts have been made to develop and codify a set of standard techniques for the assessment…
Guido Schryen
In high performance computing environments, we observe an ongoing increase in the available number of cores. For example, the current TOP500 list reveals that nine clusters have more than 1 million cores. This development calls for re-emphasizing performance (scalability) analysis and speedup laws as suggested in the…
Pablo García‐Risueño, Pablo Ibáñez
The increase of existing computational capabilities has made simulation emerge as a third discipline of Science, lying midway between experimental and purely theoretical branches [1, 2]. Simulation enables the evaluation of quantities which otherwise would not be accessible, helps to improve experiments and provides…
Gregory Valiant
We describe two algorithms for multiplying n × n matrices using time and energy O˜(n 2 ) under basic models of classical physics. The first algorithm is for multiplying integer-valued matrices, and the second, quite different algorithm, is for Boolean matrix multiplication. We hope this work inspires a deeper…
Michael Konopik, Till Korten, Eric Lutz, Heiner Linke
The fundamental energy cost of irreversible computing is given by the Landauer bound of kT ln 2 /bit. However, this limit is only achievable for infinite-time processes. We here determine the fundamental energy cost of finite-time irreversible computing within the framework of nonequilibrium thermodynamics. Comparing…
János Végh
—The paper explains why Amdahl's Law shall be interpreted specifically for distributed parallel systems and why it generated so many debates, discussions, and abuses. We set up a general model and list many of the terms affecting parallel processing. We scrutinize the validity of neglecting certain terms in different…
Nikzad Babaii Rizvandi
In recent years, the issue of energy consumption in high performance computing (HPC) systems has attracted a great deal of attention. In response to this, many energy-aware algorithms have been developed in different layers of HPC systems, including the hardware layer, service layer and system layer. These algorithms…
Leonid Yavits, Amir Morad, Uri Weiser, Ran Ginosar
— Future multiprocessor chips will integrate many different units, each tailored to a specific computation. When designing such a system, the chip architect must decide how to distribute limited system resources such as area, power, and energy among the computational units. We extend MultiAmdahl, an analytical…
Martin Karp, Niclas Jansson, Philipp Schlatter, Stefano Markidis
As supercomputers' complexity has grown, the traditional boundaries between processor, memory, network, and accelerators have blurred, making a homogeneous computer model, in which the overall computer system is modeled as a continuous medium with homogeneously distributed computational power, memory, and data movement…