21 papers · ranked by Valyu relevance
Muhammad Yasir, Li Chen, Amna Khatoon, Muhammad Amir Malik + 1 more
'Fazeel Abid'] Mixed script identification is a hindrance for automated natural language processing systems. Mixing cursive scripts of different languages is a challenge because NLP methods like POS tagging and word sense disambiguation suffer from noisy text. This study tackles the challenge of mixed script…
Analay Perez, Rae Sakakibara, Srikar Baireddy, Michael D Fetters + 2 more
'Timothy C Guetterman' 'Amaryllis Mavragani'] Title: Abstract Background The patient-physician dyad involves both verbal and nonverbal communication. Traditional methods use quantitative or qualitative coding when analyzing dyadic data of nonverbal communication. Quantitative coding methods can capture the frequency of…
Manik Sheokand, Parth Sawant
Large Language Models (LLMs) have achieved remarkable success in code generation tasks, powering various applications like code completion, debugging, and programming assistance. However, existing benchmarks such as HumanEval, MBPP, and BigCodeBench primarily evaluate LLMs on English-only prompts, overlooking the…
Dhiraj Amin, Sharvari Govilkar, Sagar Kulkarni, Yash Shashikant Lalit + 3 more
'Yash Shashikant Lalit' 'Arshi Ajaz Khwaja' 'Daries Xavier' 'Sahil Girijashankar Gupta'] Abstract— Code-mixing, the blending of linguistic elements from distinct languages to form meaningful sentences, is common in multilingual settings, yielding hybrid languages like Hinglish and Minglish. Marathi, India's third most…
Ahmad Fathan Hidayatullah, Rosyzie Anna Apong, Daphne T.C. Lai, Atika Qazi + 1 more
'Atika Qazi' 'Lexing Xie'] With the massive use of social media today, mixing between languages in social media text is prevalent. In linguistics, the phenomenon of mixing languages is known as code-mixing. The prevalence of code-mixing exposes various concerns and challenges in natural language processing (NLP)…
Baoxing PU
For a special class of three unicast sessions, in which the maximum flow from each sender to each receiver is the same positive integer k, a network coding approach is proposed. A multigeneration mixed strategy, in which (2 × n + 1) consecutive generations are taken as a mixed set, is adopted. The precoding strategy is…
Anubhav Gupta, A. Bhogal, Kripabandhu Ghosh
of Code-Mixed Sentences Authors: ['Anubhav Gupta' 'A. Bhogal' 'Kripabandhu Ghosh'] Code-mixing, the practice of alternating between two or more languages in an utterance, is a common phenomenon in multilingual communities. Due to the colloquial nature of code-mixing, there is no singular correct way to translate an…
Prashant Kodali, Anmol Goel, Likhith Asapu, Vamshi Krishna Bonagiri + 4 more
Code-Mixed Sentences Authors: ['Prashant Kodali' 'Anmol Goel' 'Likhith Asapu' 'Vamshi Krishna Bonagiri' 'Anirudh Govil' 'Monojit Choudhury' 'Manish Shrivastava' 'Ponnurangam Kumaraguru'] Current computational approaches for analysing or generating code-mixed sentences do not explicitly model "naturalness" or…
Na Wang, Yuliang Huang, Sian-Jheng Lin
The necessity of radix conversion of numeric data is an indispensable component in any complete analysis of digital computation. In this paper, we propose a binary encoding for mixed-radix digits. Second, a variant of rANS coding based on this conversion is given, which supports parallel decoding. The simulations show…
Esteban Félez Martínez, Filippo Costa, Debora Ledergerber, Lukas Imbach + 2 more
Composing individual memory traces into unified representations is fundamental to encoding of structured relationships and flexible cognition. A central debate in neuroscience concerns the neural mechanisms of these compositions: are these compositions encoded through mixed selectivity, where the same neurons…
Pauline G. Mouawad, Shievanie Sabesan, Alinka E. Greasley, Nicholas A. Lesica
The rich experience of listening to music depends on the neural integration of its constituent elements within the early auditory pathway. Here, we performed the first large-scale study of neural responses to complex music to characterize the neural coding of individual instruments and mixtures with both normal hearing…
Thomas Lynn, Julio Ottino, Richard Lueptow, Paul Umbanhowar
Cut-and-shuffle mixing is an instructive candidate system with which to assess the potential of machine learning (ML) as an approach to solve difficult mixing problems. We focus on a specific subset of cut-and-shuffle systems, the one-dimensional interval exchange transform. This class of mixing operations is well…
Zhengkai Niu, Zilong Li, Yunxiao Ma, Keke Yu + 2 more
'Mark Burke'] As bilingual families increase, the phenomenon of language mixing among children in mixed-language environments has gradually attracted academic attention. This study aims to explore the impact of language mixing on vocabulary acquisition in bilingual children and whether language distance moderates this…
Benjamin Sobkowiak, Patrick Cudahy, Melanie H. Chitwood, Taane G. Clark + 5 more
Mixed infection with multiple strains of the same pathogen in a single host can present clinical and analytical challenges. Whole genome sequence (WGS) data can identify signals of multiple strains in samples, though the precision of previous methods can be improved. Here, we present MixInfect2, a new tool to…
VP Brintha, Manikandan Narayanan
Multi-drug resistant or hetero-resistant Tuberculosis (TB) hinders the successful treatment of TB. Hetero-resistant TB occurs when multiple strains of the TB-causing bacterium with varying degrees of drug susceptibility are present in an individual. Existing studies predicting the proportion and identity of strains in…
Authors not listed
Experimental design plays an important role in efficiently acquiring informative data for system characterization and deriving robust conclusions under resource limitations. Recent advancements in high-throughput experimentation coupled with machine learning have notably improved experimental procedures. While Bayesian…
Zhangchen Xu, Yang Liu, Yueqin Yin, Mingyuan Zhou + 1 more
We introduce KODCODE, a synthetic dataset that addresses the persistent challenge of acquiring high-quality, verifiable training data across diverse difficulties and domains for training Large Language Models for coding. Existing code-focused resources typically fail to ensure either the breadth of coverage (e.g.…
Lisa Nkatha Micheni, Serawit Deyno, Joel Bazira
Background Sub-Saharan Africa, is a region that records high rates of TB infection. Mycobacterium tuberculosis mixed strain infection, especially when the strains involved are of different susceptibilities, is an area of great interest because it is linked with an increased risk of treatment failure and transmission of…
Authors not listed
Gibbs’ paradox—the apparent discontinuity in mixing entropy for gases of varying similarity and the seeming reversibility of mixing-separation cycles—has resisted fully satisfactory resolution for 150 years. We present a solution based on categorical state theory, which posits that physical con f igurations are…
Zhenlong Dai, Chang Yao, WenKang Han, Ying Yuan + 2 more
Implicit Style Representation Learning Authors: ['Zhenlong Dai' 'Chang Yao' 'WenKang Han' 'Ying Yuan' 'Zhipeng Gao' 'Jingyuan Chen'] Large Language Models (LLMs) have demonstrated great potential for assisting developers in their daily development. However, most research focuses on generating correct code, how to use…
Pablo Quijano Velasco, Kedar Hippalgaonkar, Balamurugan Ramalingam
The discovery of optimal conditions of chemical reactions is a labor-intensive, time-consuming task that requires exploring a high-dimensional parametric space. Historically the optimization of chemical reactions has been performed by manual experimentation guided by human intuition and Design of Experiments where one…