11 papers · ranked by Valyu relevance
Shahroudi, Novin, Komisarenko, Viacheslav + 2 more
Every prediction is ultimately used in a downstream task. Consequently, evaluating prediction quality is more meaningful when considered in the context of its downstream use. Metrics based solely on predictive performance often diverge from measures of real-world downstream impact. Existing approaches incorporate the…
Andrew Konya, Deger Turan, Aviv Ovadya, Lina Qui + 5 more
'Flynn Devine' 'Lisa Schirch' 'Isabella Roberts' 'Deliberative Alignment Forum'] For humanity to maintain and expand its agency into the future, the most powerful systems we create must be those which act to align the future with the will of humanity. The most powerful systems today are massive institutions like…
Mintu Dutta, Ritesh Vyas, Mohendra Roy
Self-supervised learning has emerged as a major technique for the task of learning from unlabeled data, where the current methods mostly revolve around alignment of representations and input recon struction. Although such approaches have demonstrated excellent performance in practice, their scope remains mostly…
Yu Gui, Ying Jin, Zhimei Ren
Guarantees Authors: ['Yu Gui' 'Ying Jin' 'Zhimei Ren'] Before deploying outputs from foundation models in high-stakes tasks, it is imperative to ensure that they align with human values. For instance, in radiology report generation, reports generated by a vision-language model must align with human evaluations before…
Evan Hubinger, Adam S. Jermyn, Johannes Treutlein, Rubi Hudson + 1 more
'Kate Woolverton'] Unfortunately, such approaches also raise a variety of potentially fatal safety problems, particularly surrounding situations where predictive models predict the output of other AI systems, potentially unbeknownst to us. There are numerous potential solutions to such problems, however, primarily via…
Wanjin Feng, Yuan Yuan, Jingtao Ding, Yongbo Li
In the era of increasingly complex AI models for time series forecasting, progress is often measured by marginal improvements on benchmark leaderboards. However, this approach suffers from a fundamental flaw: standard evaluation metrics conflate a model's performance with the data's intrinsic unpredictability. To…
Julian Rodemann, Esteban Garces Arias, Christoph Luther, Christoph Jansen + 1 more
'Christoph Jansen' 'Thomas Augustin'] Empirical human–AI alignment aims to make AI systems act in line with observed human behavior. While noble in its goals, we argue that empirical alignment can inadvertently introduce statistical biases that warrant caution. This position paper thus advocates against naive empirical…
Clara Fannjiang, Jennifer Listgarten
Machine learning-based design has gained traction in the sciences, most notably in the design of small molecules, materials, and proteins, with societal implications spanning drug development and manufacturing, plastic degradation, and carbon sequestration. When designing objects to achieve novel property values with…
David Medina-Ortiz, Sebastián Contreras, Juan Amado-Hinojosa, Jorge Torres-Almonacid + 3 more
'Jorge Torres-Almonacid' 'Juan A. Asenjo' 'Marcelo A. Navarrete' 'Álvaro Olivera‐Nappa'] Predicting the effect of mutations in proteins is one of the most critical challenges in protein engineering; by knowing the effect a substitution of one (or several) residues in the protein's sequence has on its overall…
Z. Shreif, Deborah A. Striegel, Vipul Periwal
A nucleotide sequence 35 base pairs long can take 1,180,591,620,717,411,303,424 possible values. An example of systems biology datasets, protein binding microarrays, contain activity data from about 40000 such sequences. The discrepancy between the number of possible configurations and the available activities is…
Charih, François, Green, James R. + 2 more
Aberrant protein-protein interactions (PPIs) underpin a plethora of human diseases, and disruption of these harmful interactions constitute a compelling treatment avenue. Advances in computational approaches to PPI prediction have closely followed progress in deep learning and natural language processing. In this…