25 papers · ranked by Valyu relevance
Arthur Francisco Araújo Fernandes, João Ricardo Rebouças Dórea, Guilherme Jordão de Magalhães Rosa
'Guilherme Jordão de Magalhães Rosa'] Computer Vision, Digital Image Processing, and Digital Image Analysis can be viewed as an amalgam of terms that very often are used to describe similar processes. Most of this confusion arises because these are interconnected fields that emerged with the development of digital…
Ali Borji
—A negative result is when the outcome of an experiment or a model is not what is expected or when a hypothesis does not hold. Despite being often overlooked in the scientific community, negative results are results and they carry value. While this topic has been extensively discussed in other fields such as social…
Frank Nielsen
A well-known old adage says that "A picture is worth a thousand words!" (attributed to the Chinese philosopher Confucius ca 500 years BC). But more precisely, what do we mean by information in images? And how can it be retrieved effectively by machines? We briefly highlight these puzzling questions in this column. But…
Haroon Idrees, Mubarak Shah, Ray Surette
Police Surveillance of Public Spaces. Historically called the "stake-out", police surveillance has a long history and evidence gained from surveillance has been an important part of investigations for nearly two centuries (Marx, 1988). Similarly, the use of visual technology by police began in the nineteenth century…
Daniel Schmid, Christian Jarvers, Heiko Neumann
Advanced computer vision mechanisms have been inspired by neuroscientific findings. However, with the focus on improving benchmark achievements, technical solutions have been shaped by application and engineering constraints. This includes the training of neural networks which led to the development of feature…
Esma Dilek, Murat Dener, Sylvain Girard
As technology continues to develop, computer vision (CV) applications are becoming increasingly widespread in the intelligent transportation systems (ITS) context. These applications are developed to improve the efficiency of transportation systems, increase their level of intelligence, and enhance traffic safety.…
Emily L. Spratt, Ahmed Elgammal
In part one of the Critique of Judgment, Immanuel Kant wrote that "the judgment of taste . . . is not a cognitive judgment, and so not logical, but is aesthetic [1]." While the condition of aesthetic discernment has long been the subject of philosophical discourse, the role of the arbiters of that judgment has more…
Seyed‐Mahdi Khaligh‐Razavi
Models of object vision have been of great interest in computer vision and visual neuroscience. During the last decades, several models have been developed to extract visual features from images for object recognition tasks. Some of these were inspired by the hierarchical structure of primate visual system, and some…
Dawei Zhang, Tingting Yang
Eye tracking is currently a research hotspot in the territory of service robotics. There is an urgent need for machine vision technique in the territory of video surveillance, and biological visual object following is one of the important basic research problems. By tracking the object of interest and recording the…
Alexander E. Siemenn, Eunice Aissi, Fang Sheng, Armi Tiihonen + 3 more
In materials research, the task of characterizing hundreds of different materials traditionally requires equally many human hours spent measuring samples one by one. We demonstrate that with the integration of computer vision into this material research workflow, many of these tasks can be automated, significantly…
Emmanuel Daucé, Pierre Albiges, Laurent Perrinet
Visual search involves a dual task of localizing and categorizing an object in the visual field of view. We develop a visuo-motor model that implements visual search as a focal accuracy-seeking policy, and we assume that the target position and category are random variables which are independently drawn from a common…
Julio Vega, Eduardo Perdices, José M. Cañas
Cameras are one of the most relevant sensors in autonomous robots. However, two of their challenges are to extract useful information from captured images, and to manage the small field of view of regular cameras. This paper proposes implementing a dynamic visual memory to store the information gathered from a moving…
Authors not listed
In experimental chemistry, actions are adjusted based on what we see—such as dosing until dissolution, heating until melting, or stirring until mixing is complete. However, current self-driving labs (SDLs) do not monitor these visual cues. HeinSight 4.0 fills this gap by integrating computer vision into SDLs to enable…
Matteo Dunnhofer, Jean de dieu Uwisengeyimana, Kohitij Kar
How does motion contribute to robust object perception when appearance cues are unreliable? In natural scenes, camouflage, clutter, and occlusion can obscure object boundaries in static images, yet humans often resolve these ambiguities once objects move. Here we ask whether modern artificial vision systems capture…
Feng Yanmin, Chen Hanlong, Bai Xue, Chen Yuanyuan + 2 more
Computer vision technology plays an important role in screening and culturing cells. This paper proposes a method to construct a helper cell library based on cell image segmentation and screening. Firstly, cell culture and image acquisition were carried out. The main content is to use laboratory conditions to carry out…
Paul Linton, Michael J. Morgan, Jenny C. A. Read, Dhanraj Vishwanath + 2 more
'Sarah H. Creem-Regehr' 'Fulvio Domini'] New approaches to 3D vision are enabling new advances in artificial intelligence and autonomous vehicles, a better understanding of how animals navigate the 3D world, and new insights into human perception in virtual and augmented reality. Whilst traditional approaches to 3D…
Jonas Kubilius, Johan Wagemans, Hans P. Op de Beeck
If a picture is worth a thousand words, as an English idiom goes, what should those words-or, rather, descriptors-capture? What format of image representation would be sufficiently rich if we were to reconstruct the essence of images from their descriptors? In this paper, we set out to develop a conceptual framework…
Niah Holtz, Evan Lloyd, Chloe Hoff, Alex C. Keene + 1 more
Cichlid fishes have long been a popular model for ecology and evolutionary biology. Not only do they exhibit nearly unparalleled taxonomic diversity, but their community structures are both complex and extremely dense. Niche partitioning along dietary axes has been credited in maintaining cichlid biodiversity in a…
Authors not listed
This work focuses on a novel human-centered digital assistant combining Mixed Reality (MR), Computer Vision and Machine Learning regression to guide professionals and students on how to operate and correctly parameterize battery manufacturing machinery. Our Concept article aim is to provide a proof of concept of our…
Malachy Guzman, Brian Geuther, Gautam Sabnis, Vivek Kumar
Changes in body mass are a key indicator of health and disease in humans and model organisms. Animal body mass is routinely monitored in husbandry and preclinical studies. In rodent studies, the current best method requires manually weighing the animal on a balance which has at least two consequences. First, direct…
Adrien Doerig, Lynn Schmittwilken, Bilge Sayim, Mauro Manassi + 1 more
Classically, visual processing is described as a cascade of local feedforward computations. Feedforward Convolutional Neural Networks (ffCNNs) have shown how powerful such models can be. Previously, using visual crowding as a well-controlled challenge, we showed that no classic model of vision, including ffCNNs, can…
Michael Chimento, Alex Hoi Hang Chan, Lucy M. Aplin, Fumihiro Kano
Collection of large behavioral data-sets on wild animals in natural habitats is vital in ecology and evolution studies. Recent progress in machine learning and computer vision, combined with inexpensive microcomputers, have unlocked a new frontier of fine-scale markerless measurements. Here, we leverage these…
Authors not listed
Determining complete atomic structures directly from microscopy images remains a longstanding challenge in materials science. MicroscopyGPT is a vision-language model (VLM) that leverages multimodal generative pre-trained transformers to predict full atomic configurations including lattice parameters, element types…
Emanuel Diamant
Traditional image processing is a field of science and technology developed to facilitate humancentered image management. But today, when huge volumes of visual data inundate our surroundings (due to the explosive growth of image-capturing devices, proliferation of Internet communication means and video sharing…
Authors not listed
Predicting protein-ligand binding affinity from three-dimensional (3D) structural data is a central task in structure-based drug discovery, yet it remains challenging due to limited data availability, structural complexity, and the sparse nature of 3D molecular representations. In this study, we investigate the…