15 papers · ranked by Valyu relevance
Ha-Myung Park, Namyong Park, Sung-Hyon Myaeng, U Kang + 1 more
'Tatsuro Kawamoto'] A connected component in a graph is a set of nodes linked to each other by paths. The problem of finding connected components has been applied to diverse graph analysis tasks such as graph partitioning, graph compression, and pattern recognition. Several distributed algorithms have been proposed to…
Pooja Yadav, Sriniwas Pandey, Sraban Kumar Mohanty
Clustering is an unsupervised learning technique in which data or objects are grouped into sets based on some similarity measure. Most of the clustering algorithms assume that the main memory is infinite and can accommodate the set of patterns. In reality many applications give rise to a large set of patterns which…
Roumaissa Ghlib, Rania Bouhadouza, Faicel Hnaien
Designing compact and efficient quantum circuits that are compatible with Noisy Intermediate-Scale Quantum (NISQ) hardware remains a central challenge in quantum computing. Most existing optimization approaches rely on fidelity-based fitness functions that require computing the full unitary matrix of the circuit.…
Alexander Ulanov, Andrey Simanovsky, Manish Marwah
—Present day machine learning is computationally intensive and processes large amounts of data. It is implemented in a distributed fashion in order to address these scalability issues. The work is parallelized across a number of computing nodes. It is usually hard to estimate in advance how many nodes to use for a…
Hassan Mushtaq, Sajid Gul Khawaja, Muhammad Usman Akram, Amanullah Yasin + 3 more
Clustering is the most common method for organizing unlabeled data into its natural groups (called clusters), based on similarity (in some sense or another) among data objects. The Partitioning Around Medoids (PAM) algorithm belongs to the partitioning-based methods of clustering widely used for objects categorization…
Yun Xu, Wenhua Cheng, Pengyu Nie, Fengfeng Zhou + 1 more
Haplotype phasing represents an essential step in studying the association of genomic polymorphisms with complex genetic diseases, and in determining targets for drug designing. In recent years, huge amounts of genotype data are produced from the rapidly evolving high-throughput sequencing technologies, and the data…
Anuj Sharma, Syed Mohammed Arshad Zaidi
Graphs and their traversal is becoming significant as it is applicable to various areas of mathematics, science and technology. Various problems in fields as varied as biochemistry (genomics), electrical engineering (communication networks), computer science (algorithms and computation) can be modeled as Graph…
Michael Bar-Sinai
Storing and manipulating Big Data relies on various data structures, algorithms and technologies. Some of these are new, while others have existed for quite a while (the Bloom filter was presented in 1970) and are now making their way into mainstream software engineering. We present those algorithms and technologies…
Hsiang-Huang Wu, Chien‐Min Wang, Hsuan-Chi Kuo, Wei-Chun Chung + 1 more
'Jan-Ming Ho'] Abstract—Suffix Array (SA) is a cardinal data structure in many pattern matching applications, including data compression, plagiarism detection and sequence alignment. However, as the volumes of data increase abruptly, the construction of SA is not amenable to the current large-scale data processing…
Muhammad Idris, Shujaat Hussain, Muhammad Hameed Siddiqi, Waseem Hassan + 3 more
'Waseem Hassan' 'Hafiz Syed Muhammad Bilal' 'Sungyoung Lee' 'Christophe Antoniewski'] Large quantities of data have been generated from multiple sources at exponential rates in the last few years. These data are generated at high velocity as real time and streaming data in variety of formats. These characteristics give…
Saurav Prakash, Amirhossein Reisizadeh, Ramtin Pedarsani, Salman Avestimehr
'Salman Avestimehr'] To combat the growing demands for efficient processing of large scale graph-structured datasets, many distributed graph computing systems have been developed recently. As these systems require many messages to be exchanged among computing machines at each step of the computation, communication…
Manuel Penschuck
Shuffling is the process of rearranging a sequence of elements into a random order such that any permutation occurs with equal probability. It is an important building block in a plethora of techniques used in virtually all scientific areas. Consequently considerable work has been devoted to the design and…
Silu Huang, Ada Wai-Chee Fu
As computer clusters are found to be highly effective for handling massive datasets, the design of efficient parallel algorithms for such a computing model is of great interest. We consider (α, k)-minimal algorithms for such a purpose, where α is the number of rounds in the algorithm, and k is a bound on the deviation…
Ricardo Villanueva-Polanco
In this paper, we will study the key enumeration problem, which is connected to the key recovery problem posed in the cold boot attack setting. In this setting, an attacker with physical access to a computer may obtain noisy data of a cryptographic secret key of a cryptographic scheme from main memory via this data…
Martin Werner
This paper provides an abstract analysis of parallel processing strategies for spatial and spatio-temporal data. It isolates aspects such as data locality and computational locality as well as redundancy and locally sequential access as central elements of parallel algorithm design for spatial data. Furthermore, the…