15 papers · ranked by Valyu relevance
Ting Yao, Yiheng Zhang, Zhaofan Qiu, Yingwei Pan + 1 more
A steady momentum of innovations and breakthroughs has convincingly pushed the limits of unsupervised image representation learning. Compared to static 2D images, video has one more dimension (time). The inherent supervision existing in such sequential structure offers a fertile ground for building unsupervised…
Gabriele Berton, Gabriele Trivigno, Barbara Caputo, Carlo Masone
—Visual Place Recognition aims at recognizing previously visited places by relying on visual clues, and it is used in robotics applications for SLAM and localization. Since typically a mobile robot has access to a continuous stream of frames, this task is naturally cast as a sequence-to-sequence localization problem.…
Viacheslav Osaulenko
In this paper we start with a simple question, how is it possible that humans can recognize different movements over skin with only a prior visual experience of them? Or in general, what is the representation of spatial sequences that are invariant to scale, rotation, and translation across different modalities? To…
Vincent Michalski, Roland Memisevic, Kishore Konda
Bi-linear feature learning models, like the gated autoencoder, were proposed as a way to model relationships between frames in a video. By minimizing reconstruction error of one frame, given the previous frame, these models learn "mapping units" that encode the transformations inherent in a sequence, and thereby learn…
Riccardo Mereu, Gabriele Trivigno, Gabriele Berton, Carlo Masone + 1 more
'Barbara Caputo'] Abstract— In robotics, Visual Place Recognition is a continuous process that receives as input a video stream to produce a hypothesis of the robot's current position within a map of known places. This task requires robust, scalable, and efficient techniques for real applications. This work proposes a…
Sourav Garg, Michael Milford
©2021 IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses, in any current or future media, including reprinting/republishing this material for advertising or promotional purposes, creating new collective works, for resale or redistribution to servers or lists, or…
Qiong Liu, Yanxia Zhang
—As data from IoT (Internet of Things) sensors become ubiquitous, state-of-the-art machine learning algorithms face many challenges on directly using sensor data. To overcome these challenges, methods must be designed to learn directly from sensors without manual annotations. This paper introduces Sensory Time-cue for…
Yuwei Cui, Subutai Ahmad, Jeff Hawkins
The ability to recognize and predict temporal sequences of sensory inputs is vital for survival in natural environments. Based on many known properties of cortical neurons, hierarchical temporal memory (HTM) sequence memory is recently proposed as a theoretical framework for sequence learning in the cortex. In this…
Athanasios Efthymiou, Stevan Rudinac, Monika Kackovic, Nachoem M. Wijnberg + 1 more
Artistic Sequences Authors: ['Athanasios Efthymiou' 'Stevan Rudinac' 'Monika Kackovic' 'Nachoem M. Wijnberg' 'Marcel Worring'] We propose Set2Seq Transformer, a novel sequential multiple instance architecture, that learns to rank permutation aware set representations of sequences. First, we illustrate that learning…
Isma Hadji, Konstantinos G. Derpanis, Allan D. Jepson
We introduce a weakly supervised method for representation learning based on aligning temporal sequences (e.g., videos) of the same process (e.g., human action). The main idea is to use the global temporal ordering of latent correspondences across sequence pairs as a supervisory signal. In particular, we propose a loss…
Jayanta Dey, Soures, Nicholas, Miranda Gonzales + 3 more
In this pilot study, we propose a neuro-inspired approach that compresses temporal sequences into context-tagged chunks, where each tag represents a recurring structural unit or "community" in the sequence. These tags are generated during an offline sleep phase and serve as compact references to past experience…
M. Zhao, Chengxu Zhuang, Yizhou Wang, Tai Sing Lee
We propose a new neurally-inspired model that can learn to encode the global relationship context of visual events across time and space and to use the contextual information to modulate the analysis by synthesis process in a predictive coding framework. The model learns latent contextual representations by maximizing…
Florian Feiler, Emre Neftci, Younes Bouhadjar
Networks Authors: ['Florian Feiler' 'Emre Neftci' 'Younes Bouhadjar'] Abstract—The ability to predict future events or patterns based on previous experience is crucial for many applications such as traffic control, weather forecasting, or supply chain management. While modern supervised Machine Learning approaches…
Vadym Gryshchuk, Cornelius Weber, Chu Kiong Loo, Stefan Wermter
Lifelong learning is a long-standing aim for artificial agents that act in dynamic environments, in which an agent needs to accumulate knowledge incrementally without forgetting previously learned representations. We investigate methods for learning from data produced by event cameras and compare techniques to mitigate…
Mingcan Yu, Junying Wang
Although principles of neuroscience like reinforcement learning, visual perception and attention have been applied in machine learning models, there is a huge gap between machine learning and mammalian learning. Based on the advances in neuroscience, we propose the "context sequence theory" to give a common explanation…