14 papers · ranked by Valyu relevance
Xiyin Zeng, Yuyu Sun, Haoyang Li, Shouqiang Liu + 1 more
Vision-Language-Action systems follow instructions to execute multi-step tasks in multimodal environments. Recent VLA approaches typically rely on post-hoc correction mechanisms or operate under fixed task decompositions and alignment schemes. However, once an intermediate step is mis-specified, local errors propagate…
Mintu Dutta, Ritesh Vyas, Mohendra Roy
Self-supervised learning has emerged as a major technique for the task of learning from unlabeled data, where the current methods mostly revolve around alignment of representations and input recon struction. Although such approaches have demonstrated excellent performance in practice, their scope remains mostly…
Corentin Lobet, Francesca Chiaromonte
Feature attribution is the dominant paradigm for explaining deep neural networks. However, most existing methods only loosely reflect the model's prediction-making process, thereby merely white-painting the black box. We argue that explanatory alignment is a key aspect of trustworthiness in prediction tasks…
Mohammad Asadi, Soheil Hor, Bardiya Akhbari, Jack W. O'Sullivan + 5 more
Retrieval-augmented forecasting promises to adapt frozen Time Series Foundation Models (TSFMs) to new domains without fine-tuning, but recent methods typically rely on learned fusion modules, i.e., trained adapters that merge retrieved examples into the backbone's forecast, based on the assumption that frozen backbones…
Tiejin Chen, Xiaoou Liu, Vishnu Nandam, Kuan-Ru Liou + 1 more
Preference-based alignment like Reinforcement Learning from Human Feedback (RLHF) learns from pairwise preferences, yet the labels are often noisy and inconsistent. Existing uncertainty-aware approaches weight preferences, but ignore a more fundamental factor: the reliability of the answers being compared. To address…
Moisés Santos, Peter van der Putten, Bernhard Pfahringer, Carlos Soares
We propose Rashomon Alignment (RA), a new measure to assess functional similarity between two models. Existing functional similarity measures are distributional, quantifying differences between outputs of models applied to real-world data. However, these measures can be regarded as ecologically valid only for regions…
Yuhan Huang, Huanran Chen, Yinpeng Dong
Although Large Language Models (LLMs) achieve strong alignment through supervised fine-tuning and reinforcement learning from human feedback, the alignment is often fragile under subsequent fine-tuning. Existing explanations either attribute alignment fragility to gradient geometry or characterize it as a…
Shuxuan Li, Zhilin Zhao, Quyu Kong, Wei-Shi Zheng
Performance estimation under distribution shift aims to predict how a model behaves on an unlabeled test set whose distribution differs from the training data, a scenario that requires reliable indicators that can faithfully reflect model behavior without ground-truth labels. Existing approaches rely solely on the…
Ni Yang, Rui He, Philipp Homan, Iris Sommer + 2 more
Large language models (LLMs) reliably predict neural activity during language comprehension and transformer depth has been interpreted as mirroring hierarchical cortical organization. However, it remains unclear whether such alignment extends to subcortical regions, overlaps spatially across languages, and what the…
Gal Fybish, Teo Susnjak
Machine learning models are increasingly used in high-stakes domains where their predictions can actively shape the environments in which they operate, a phenomenon known as performative prediction. This dynamic, in which the deployment of the model influences the very outcome it seeks to predict, can lead to…
Yuhong Luo, David M. Pennock, Xintong Wang
It is increasingly common to aggregate predictions from multiple LLMs, each with domain expertise or access to private tools and data, to improve collective prediction performance. In decentralized settings, aggregation weights need to be determined without access to models' private information and should remain robust…
Vasiliki Tassopoulou, Charis Stamouli, Haochang Shou, George J. Pappas + 1 more
Despite recent progress in predicting biomarker trajectories from real clinical data, uncertainty in the predictions poses high-stakes risks (e.g., misdiagnosis) that limit their clinical deployment. To enable safe and reliable use of such predictions in healthcare, we introduce a conformal method for…
Yiyao Yang
Robust machine learning for regulatory sequence modeling is examined under biologically and technically induced distribution shifts. Although deep convolutional and attention-based architectures achieve strong in-distribution performance on regulatory tasks, they are predominantly evaluated under i.i.d. assumptions…
Tobias Rønlev-Knudsen, Henrik Madsen, Jan Kloppenborg Møller
We present a framework for online and adaptive forecasting and hierarchical reconciliation using linear regression models. We begin by formalizing hierarchies using graphs, and motivated by their structure, formulate a multivariate linear model using the matrix normal distribution to characterize residuals. Parameter…