13 papers · ranked by Valyu relevance
Xing Gao, Dazhong Rong, Qinming He
Investigating the mapping between visual stimuli and neural responses in the visual cortex contributes to a deeper understanding of biological visual processing mechanisms. Most existing studies characterize this mapping by training models to directly encode visual stimuli into neural responses or decode neural…
Shuxiao Ma, Linyuan Wang, Bin Yan
Biological research has revealed that the verbal semantic information in the brain cortex, as an additional source, participates in nonverbal semantic tasks, such as visual encoding. However, previous visual encoding models did not incorporate verbal semantic information, contradicting this biological finding. This…
Shuxiao Ma, Linyuan Wang, Senbao Hou, Bin Yan
Recently, there has been a surge in the popularity of pretrained large language models (LLMs) (such as GPT-4), sweeping across the entire Natural Language Processing (NLP) and Computer Vision (CV) communities. These LLMs have demonstrated advanced multi-modal understanding capabilities and showcased strong performance…
Peiying Zhang, Chenhui Li, Changbo Wang
At a high level, the goal of the encoder network is to embed a large amount of information into a graphical chart while leaving the coded image perceptually identical to the original. A straightforward solution is to train a model to minimize the mean squared error (MSE) of the pixel difference between the original…
Khaled Masmoudi, Marc Antonini, Pierre Kornprobst
We propose the design of an original scalable image coder/decoder that is inspired from the mammalians retina. Our coder accounts for the time-dependent and also nondeterministic behavior of the actual retina. The present work brings two main contributions: As a first step, (i) we design a deterministic image coder…
Aleksandra Pižurica
This paper presents a convenient graphical tool for encoding visual patterns (such as image patches and image atoms) as point constellations in a space spanned by perceptual features and with a clear geometrical interpretation. General theory and a practical pattern encoding scheme are presented, inspired by encoding…
Søren Rasmussen, Karsten Østergaard Noe, Oliver Gyldenberg Hjermitslev, Henrik Pedersen
'Oliver Gyldenberg Hjermitslev' 'Henrik Pedersen'] Abstract—We introduce DeepMorph, an information embedding technique for vector drawings. Provided a vector drawing, such as a Scalable Vector Graphics (SVG) file, our method embeds bitstrings in the image by perturbing the drawing primitives (lines, circles, etc.).…
Li Zhaoping
Our brain recognizes only a tiny fraction of sensory input, due to an information processing bottleneck. This blinds us to most visual inputs. Since we are blind to this blindness, only a recent framework highlights this bottleneck by formulating vision as mainly looking and seeing. Looking selects a tiny fraction of…
J. Gerard Wolff
The SP theory of intelligence aims to simplify and integrate concepts in computing and cognition, with information compression as a unifying theme. This article is about how the SP theory may, with advantage, be applied to the understanding of natural vision and the development of computer vision. Potential benefits…
Anna Cattani, Gaute T. Einevoll, Stefano Panzeri
The phase-of-firing code is a neural coding scheme whereby neurons encode information using the time at which they fire spikes within a cycle of the ongoing oscillatory pattern of network activity. This coding scheme may allow neurons to use their temporal pattern of spikes to encode information that is not encoded in…
Jesús Malo
Color Appearance Models are biological networks that consist of a cascade of linear+nonlinear layers that modify the linear measurements at the retinal photo-receptors leading to an internal (nonlinear) representation of color that correlates with psychophysical experience. The basic layers of these networks include…
Rufin VanRullen, Leila Reddy
While objects from different categories can be reliably decoded from fMRI brain response patterns, it has proved more difficult to distinguish visually similar inputs, such as different instances of the same category. Here, we apply a recently developed deep learning system to the reconstruction of face images from…
Aman Chawla
In this brief paper, the authors study the tuning curves of starburst amacrine cells (SACs) and introduce a quantity called the irresolution or ambiguity of a SAC. They show that the rate of data generated by a starburst amacrine cell is inversely proportional to its irresolution. This is done by providing bounds on…