10 papers · ranked by Valyu relevance
Hamel Husain, Ho-Hsiang Wu, Tiferet Gazit, Miltiadis Allamanis + 1 more
'Marc Brockschmidt'] To enable evaluation of progress on code search, we are releasing the CodeSearchNet Corpus and are presenting the CodeSearch-Net Challenge, which consists of 99 natural language queries with about 4k expert relevance annotations of likely results from Code-SearchNet Corpus. The corpus contains…
Chen Wu, Ming Yan
Semantic code search is the task of retrieving relevant code snippet given a natural language query. Different from typical information retrieval tasks, code search requires to bridge the semantic gap between the programming language and natural language, for better describing intrinsic concepts and semantics.…
Yutao Xie, Jiayi Lin, Hande Dong, Lei Zhang + 1 more
Code writing is repetitive and predictable, inspiring us to develop various code intelligence techniques. This survey focuses on code search, that is, to retrieve code that matches a given natural language query by effectively capturing the semantic similarity between the query and code. Deep learning, being able to…
Andor Diera, Abdelhalim Hafedh Dahou, Lukas Galke, Fabian Karl + 2 more
'Florian Sihler' 'Ansgar Scherp'] Language models can serve as a valuable tool for software developers to increase productivity. Large generative models can be used for code generation and code completion, while smaller encoder-only models are capable of performing code search tasks using natural language queries.…
Ivan Sedykh, Dmitry Abulkhanov, Nikita Sorokin, Sergey Nikolenko + 1 more
'Valentin Malykh'] Code search is an important task that has seen many developments in recent years. However, previous attempts have mostly considered the problem of searching for code by a text query. We argue that using a code snippet (and possibly an associated traceback) as a query and looking for answers with…
Shushan Arakelyan, Anna Hakhverdyan, Miltiadis Allamanis, Christophe Hauser + 2 more
'Christophe Hauser' 'Luis García' 'Xiang Ren'] Semantic code search is the task of retrieving a code snippet given a textual description of its functionality. Recent work has been focused on using similarity metrics between neural embeddings of text and code. However, current language models are known to struggle with…
Zhengyu Zhao, Yuanyuan Lu, Yijie Tong, Xin Chen + 1 more
Discriminative traits are important in biodiversity and macroevolution, but extracting and representing these features from huge natural history collections using traditional methods can be challenging and time-consuming. To fully utilize the collections and their associated metadata, it is urgent now to increase the…
Marc-Antoine Jacques, Maciej Dobrzyński, Paolo Armando Gagliardi, Raphael Sznitman + 1 more
Fluorescent biosensors routinely yield thousands of single-cell, heterogeneous, multi-dimensional signaling trajectories that are difficult to mine for relevant information. We present CODEX, an approach based on artificial neural networks to guide exploration of time-series datasets and to identify motifs in dynamic…
Bohdan B. Khomtchouk, Kasra A. Vand, Thor Wahlestedt, Kelly Khomtchouk + 2 more
We propose a search engine and file retrieval system for all bioinformatics databases worldwide. PubData searches biomedical data in a user-friendly fashion similar to how PubMed searches biomedical literature. PubData is built on novel network programming, natural language processing, and artificial intelligence…
Patrick Wu, Aliya Gifford, Xiangrui Meng, Xue Li + 7 more
Many studies of Electronic Health Record (EHR) data utilize custom-developed aggregations of billing codes enabling clinical and genetic research, including phenome-wide association studies (PheWAS). One such grouping is the phecode system, originally developed for PheWAS. Phecodes were built upon the International…