23 papers · ranked by Valyu relevance
Mara Thomas, Frants H. Jensen, Baptiste Averly, Vlad Demartsev + 4 more
The manual detection, analysis, and classification of animal vocalizations in acoustic recordings is laborious and requires expert knowledge. Hence, there is a need for objective, generalizable methods that detect underlying patterns in these data, categorize sounds into distinct groups, and quantify similarities…
Dan Mi, Lu Qin
National music is a treasure of Chinese traditional culture. It contains the cultural characteristics of various regions and reflects the core value of Chinese traditional culture. Classification technology classifies a large number of unorganized drama documents, which are not labeled, and to some extent, it helps…
Li Jiashen, Zhang Xianwu, Nattapol Aunsri
In speech signal processing, time-frequency analysis is commonly employed to extract the spectrogram of speech signals. While many algorithms exist to achieve this with high-quality results, they often lack the flexibility to adjust the resolution of the extracted spectrograms. However, applications such as speech…
Johanna Devaney
Digital audio processing tools offer music researchers the opportunity to examine both non-notated music and music as performance. This chapter summarises the types of information that can be extracted from audio as well as currently available audio tools for music corpus studies. The survey of extraction methods…
Martino Trapanotto, Loris Nanni, Sheryl Brahnam, Xiang Guo + 1 more
The classification of vocal individuality for passive acoustic monitoring (PAM) and census of animals is becoming an increasingly popular area of research. Nearly all studies in this field of inquiry have relied on classic audio representations and classifiers, such as Support Vector Machines (SVMs) trained on…
Sigal Saar, Partha P. Mitra, Brian McCabe
The developmental trajectory of nervous system dynamics shows hierarchical structure on time scales spanning ten orders of magnitude from milliseconds to years. Analyzing and characterizing this structure poses significant signal processing challenges. In the context of birdsong development, we have previously proposed…
Authors not listed
Acoustic measurements of batteries are known to be correlated to their state-of-charge, creating opportunities for state estimation that do not rely on electrical signals. State estimators are typically parametric models fitted from data, often from the broad toolbox of machine learning. Such models can be easily…
Vladimir Kudriavtsev, Vladimir Polyshchuk, Douglas L Roy
A new method and application is proposed to characterize intensity and pitch of human heart sounds and murmurs. Using recorded heart sounds from the library of one of the authors, a visual map of heart sound energy was established. Both normal and abnormal heart sound recordings were studied. Representation is based on…
Shahin Tavakoli, Beatrice Matteo, Davide Pigoli, Eleanor Chodroff + 4 more
'John Coleman' 'Michele Gubian' 'Margaret E. L. Renwick' 'Morgan Sonderegger'] Phonetics is the scientific field concerned with the study of how speech is produced, heard and perceived. It abounds with data, such as acoustic speech recordings, neuroimaging data, or articulatory data. In this paper, we provide an…
Fengrong He, Ian H. Stevenson, Monty A. Escabi
Theories of efficient coding propose that the auditory system is optimized for the statistical structure of natural sounds, yet the transformations underlying optimal acoustic representations are not well understood. Using a database of natural sounds including human speech and a physiologically-inspired auditory…
Aarón López-García
Audio fingerprinting is a technique used to identify and match audio recordings based on their unique characteristics. It involves creating a condensed representation of an audio signal that can be used to quickly compare and match against other audio recordings. The fingerprinting process involves analyzing the audio…
Ama Marina Kreme, Adrien Meynard
—We present a spectrogram separation method tailored for mixtures comprising two nonstationary components. By exploiting the unique characteristics of their time-frequency representations, we propose an inverse problem formulation to estimate the spectrograms of the components. We then introduce an alternating…
Shruti Gupta, Md Shah Fahad, Akshay Deepak
—Convolutional neural networks (CNN) are widely used for speech emotion recognition (SER). In such cases, the short time fourier transform (STFT) spectrogram is the most popular choice for representing speech, which is fed as input to the CNN. However, the uncertainty principles of the short-time Fourier transform…
M Rolland, A Zai, RHR Hahnloser, C Del Negro + 1 more
Human language learning and maintenance depend primarily on auditory feedback but are also shaped by other sensory modalities. Individuals who become deaf after learning to speak (post-lingual deafness) experience a gradual decline in their language abilities. A similar process occurs in songbirds, where deafness leads…
Andrey Anikin, Christian T. Herbst
We address two research applications in this methodological review: starting from an audio recording, the goal may be to characterize nonlinear phenomena (NLP) at the level of voice production or to test their perceptual effects on listeners. A crucial prerequisite for this work is the ability to detect NLP in acoustic…
Shukai Chen, Marvin Thielk, Timothy Q. Gentner
Studies comparing acoustic signals often rely on pixel-wise differences between spectrograms, as in for example mean squared error (MSE). Pixel-wise errors are not representative of perceptual sensitivity, however, and such measures can be highly sensitive to small local signal changes that may be imperceptible. In…
Julio C. Hechavarría, M. Jerome Beetz, Francisco Garcia-Rosales, Manfred Kössl
Communication sounds are ubiquitous in the animal kingdom, where they play a role in advertising physiological states and/or socio-contextual scenarios. Distress sounds, for example, are typically uttered in distressful scenarios such as agonistic interactions. Here, we report on the occurrence of superfast temporal…
Nai Ding, Aniruddh D. Patel, Lin Chen, Henry Butler + 2 more
Speech and music have structured rhythms, but these rhythms are rarely compared empirically. This study, based on large corpora, quantitatively characterizes and compares a major acoustic correlate of spoken and musical rhythms, the slow (0.25-32 Hz) temporal modulations in sound intensity. We show that the speech…
Guido Pauli, G. Joseph Ray, Anton Bzhelyansky, Birgit Jaki + 18 more
Classical 1D 1H NMR spectra are prototypic for NMR spectroscopy in that they represent a wealth of chemical information encoded into convoluted graphs or patterns that contain complex features (aka multiplets), even for seemingly simple molecules. Accordingly, the utility of NMR depends on the theoretical and visual…
Authors not listed
This report compares various simulation and data analysis methods for free induction decay (FID) signals in Nuclear Magnetic Resonance (NMR) Spectroscopy. The methods discussed include discrete fast Fourier transformation (FFT), least squares fitting (LSF), short-time Fourier transformation (STFT), and wavelet…
Authors not listed
Nuclear magnetic resonance spectroscopy (NMR) is one of the most potent analytical chemistry methods, providing a unique insight into molecular structures. Its non-invasiveness makes it a perfect tool for monitoring chemical reactions and determining their products and kinetics. Typically, the reactions are monitored…
Amelia Carolina Sparavigna
Previous studies (Sparavigna, 2023) have demonstrated the Tsallis q-Gaussian functions suitable for the analysis of Raman spectra. These functions can be used for simulating the different lineshapes of Raman bands. Here we apply q-Gaussians to graphite Raman spectra from RRUFF and Raman Open Database.
Authors not listed
Nondestructive ultrasonic testing is finding increasing use in battery science. We provide instructions and software for the development of a low cost, modular, and easy to use scanning acoustic microscope. Basic principles of ultrasonic testing are discussed with particular attention to its application for operando…