10 papers · ranked by Valyu relevance
Matteo Cinelli, Leto Peel, Antonio Iovanella, Jean‐Charles Delvenne
We consider the network constraints on the bounds of the assortativity coefficient, which aims to quantify the tendency of nodes with the same attribute values to be connected. The assortativity coefficient can be considered as the Pearson's correlation coefficient of node metadata values across network edges and lies…
Daniel Engel, Freek Verbeek, Pranav Kumar, Binoy Ravindran
The binary executable format is the standard method for distributing and executing software. Yet, it is also as opaque a representation of software as can be. If the binary format were augmented with metadata that provides security-relevant information, such as which data is intended by the compiler to be executable…
Dae Il Kim, Michael C. Hughes, Erik B. Sudderth
We introduce the nonparametric metadata dependent relational (NMDR) model, a Bayesian nonparametric stochastic block model for network data. The NMDR allows the entities associated with each node to have mixed membership in an unbounded collection of latent communities. Learned regression models allow these memberships…
Basavaraja Bheemalingappa Sagar, Georg von Hippel, Giannis Koutsou, Hideo Matsufuru + 3 more
- 𝑑Computing Research Center, High Energy Accelerator Research Organization (KEK), and Accelerator Science Program, Graduate Institute for Advanced Studies, Graduate University for Advanced Studies (SOKENDAI), Oho 1-1, Tsukuba 305-0801, Japan - 𝑒Bernoulli Institute for Mathematics, Computer Science and Artificial…
Peng Sun, Yonggang Wen, Duong Nguyen Binh Ta, Haiyong Xie
—In large-scale distributed file systems, efficient metadata operations are critical since most file operations have to interact with metadata servers first. In existing distributed hash table (DHT) based metadata management systems, the lookup service could be a performance bottleneck due to its significant CPU…
Salman Niazi, Mahmoud Ismail, Seif Haridi, Jim Dowling
Recent improvements in both the performance and scalability of shared-nothing, transactional, in-memory NewSQL databases have reopened the research question of whether distributed metadata for hierarchical file systems can be managed using commodity databases. In this paper, we introduce HopsFS, a next generation…
Authors not listed
—Distributed storage architectures are foundational to modern cloud-native infrastructure, yet a critical operational bottleneck persists within disaster recovery (DR) workflows: the dependence on content-based cryptographic hashing for data identification and synchronization. While hash-based deduplication is…
Lena Mangold, Camille Roth
Network analysis is often enriched by including an examination of node metadata. In the context of understanding the mesoscale of networks it is often assumed that node groups based on metadata and node groups based on connectivity patterns are intrinsically linked. This assumption is increasingly being challenged…
Spyros Blanas, Suren Byna
Advances in technology and computing hardware are enabling scientists from all areas of science to produce massive amounts of data using large-scale simulations or observational facilities. In this era of data deluge, effective coordination between the data production and the analysis phases hinges on the availability…
Thurston H. Y. Dang, José Cambronero, Martin Rinard
We present BIEBER (Byte-IdEntical Binary parsER), the first system to model and regenerate a full working parser from instrumented program executions. To achieve this, BIEBER exploits the regularity (e.g., header fields and array-like data structures) that is commonly found in file formats. Key generalization steps…