Volume A · since 1991
arXiv
The preprint server that rewired how physics, maths and machine learning are published.
Physics · Maths · Computer Science
Newest strong work in arXiv
Volume A · since 1991
The preprint server that rewired how physics, maths and machine learning are published.
Physics · Maths · Computer Science
Newest strong work in arXiv
Rishabh Shukla, Adithya Santhosh, Shaili Gandhi, Samrudh Moode + 1 more
Diffusion policies have shown strong performance in learning complex, multi-modal behaviors for robotic manipulation. However, their application to contact-rich disassembly tasks remains limited by a key trade-off: the iterative denoising process introduces inference latencies that makes high frequency control…
Dor Elimelech, Victor V. Albert, Alexander Barg
We develop a theory of approximate quantum error correction (QEC) based on the error-set model, complemented by general methods for code construction. Exact QEC has a powerful error-set structure: by the Knill-Laflamme conditions, a code correcting a given error set automatically protects against every channel whose…
Yunyao Zhang, Xinglang Zhang, Zeliang Chen, Junqing Yu + 1 more
Large language models (LLMs) have become powerful tools for language understanding and logical reasoning. However, they still make mistakes when a problem requires both understanding meaning and following logic. A key reason is that natural-language statements often carry implicit semantic relations before any formal…
Ian Whitehouse, Anıl Zenginoğlu, Franz Klein, Mohe Edeen Abu Maizer + 4 more
We draw a structural analogy between quantum error correction (QEC) and error handling in neural circuits with respect to their redundant encodings and constraint-based inferences. In QEC, logical information is embedded in a protected codespace within a larger Hilbert space. A set of commuting checks (e.g. stabilizer…
Jeong Min Kong, Richard S. Sutton
Neural networks that can grow or both grow and shrink during learning, referred to as growing neural networks and elastic neural networks, respectively, have recently been explored in offline continual learning with a particular focus on catastrophic forgetting. Driven by the observations that 1) online continual…
Xinyi Zheng, Ling Shi, Tianlong Yu, Yongxin Zhao + 2 more
Large Language Models (LLMs) have made significant progress in reasoning, particularly in deductive reasoning, which is crucial for high-stakes decision-making. As models improve, evaluation benchmarks should evolve to keep pace. However, existing benchmarks lack fine-grained control over logical complexity and…
Zijian Zhu, Menglin Zou, Zhuang Li, Yaojie Tu + 1 more
Vision-Language-Action (VLA) models have emerged as a promising paradigm for general-purpose robot control. However, their performance remains fundamentally constrained by the availability of high-quality robot trajectory data. In current robot learning practice, such data are primarily collected through human…
Xiandong Zou, Jing Huang, Jianshu Li, Pan Zhou
Large Language Models (LLMs) increasingly rely on intermediate reasoning, yet explicit Chain-of-Thought (CoT) suffers from a linguistic space bottleneck: each thought must be decoded into tokens, causing high inference overhead. Latent reasoning moves deliberation into continuous space, but existing methods mostly…
Wolfgang Pietsch
Because large language models (LLMs) are impressively successful in predicting text, it appears that they must have access to a 'world model' representing causal and definitional structure. However, the dominant formalisms of modern causal inference -- Judea Pearl's interventionist approach and the Neyman-Rubin…
Leslie G. Valiant
In current Large Language Models we can trust the production of smoothly flowing prose on the basis of the principles of machine learning. However, there is no comparably principled basis to justify trust in the content of the text produced. It appears to be conventional wisdom that addressing this issue by adding more…
Jiawei Gao, Chaoqi Liu, Peilin Wu, Haonan Chen + 1 more
Real-world robotic manipulation tasks often involve forceful interactions with the environment, such as using tools of varying weights, transporting objects with different masses, and performing contact-rich tasks like table wiping. Previous learning-based approaches typically employ imitation learning policies that…
Csaba Czabán, Orsolya Kálmán, Sergey N. Filippov, Zoltán Zimborás
Readout errors are one of the dominant sources of noise in current quantum processors, limiting both expectation-value estimation and sampling-based applications. Since they affect only the classical measurement outcomes, they can be addressed using classical coding techniques: immediately before measurement, each data…
Harsh Gupta, Guanya Shi, Wenzhen Yuan
The most widely-adopted robot learning pipelines today learn skills from robot demonstrations or structured human data, which are expensive to collect and tied to specific embodiments. In contrast, unstructured human videos provide a scalable alternative. They contain diverse manipulation demonstrations across objects…
Yiwen Qiu, Linjuan Wu, Yizhou Liu, Yuchen Yan + 8 more
Large language models have achieved remarkable progress on complex reasoning tasks. However, they often implicitly fabricate information when inputs are incomplete, producing confident but unreliable conclusions -- a failure mode we term ungrounded reasoning. We argue that this issue arises not from insufficient…
Marina Igitkhanian, Erik Arakelyan
Recently, language models have made rapid progress across various domains and applications. However, their capability for self-improvement, i.e., whether they are adept at recognising and correcting flaws in their own reasoning, remains dubious. In this study, we address this question by constructing a sufficiency test…
Shaked Regev, Daniel Dilley, Andrea Delgado, Ryan Bennink
We propose a novel method to calculate logical error rates in surface codes, assuming independent and identically distributed physical errors. We show how to use our method to analyze hypothetical quantum computers with various configurations and select designs with lower error rates. Currently, this requires expensive…
Zenghui Zhou, Man Li, Xiaoke Fang, Xinyi Zhou + 2 more
Large Language Models (LLMs) achieve strong performance on logical reasoning benchmarks, yet their reliability remains uncertain. Existing evaluations rely on static benchmarks, which fail to assess robustness under logically equivalent transformations and often overestimate reasoning capability. We propose LGMT…
Snehal Jauhri, Vignesh Prasad, Georgia Chalvatzaki
Mobile Manipulation (MoMa) of articulated objects, such as opening doors, drawers, and cupboards, demands simultaneous, whole-body coordination between a robot's base and arms. Classical whole-body controllers (WBCs) can solve such problems via hierarchical optimization, but require extensive hand-tuned optimization…
Xiaodie Lin, Linxuan Li, Haidong Yuan
The precision and sensitivity achievable in quantum metrology are often compromised by the presence of noise. While quantum error correction has emerged as a promising strategy, it is ineffective in addressing noise that is indistinguishable from the signal. To address this challenge, virtual state purification was…
Yinan Chen, Zongyuan Wang, Sisi Zhou
We introduce a new method for error-corrected quantum metrology where only partial quantum error correction (QEC) is needed to suppress local noise and maintain the probe states' super-standard-quantum-limit (super-SQL) sensing performance. This stands in contrast to the existing QEC-assisted sensing schemes in Phys.…