14 papers · ranked by Valyu relevance
Moustafa Alzantot, Luis García, Mani Srivastava
Generative models such as the variational autoencoder (VAE) and the generative adversarial networks (GAN) have proven to be incredibly powerful for the generation of synthetic data that preserves statistical properties and utility of realworld datasets, especially in the context of image and natural language text.…
Reinhard Heckel, Paul Hand
Deep neural networks, in particular convolutional neural networks, have become highly effective tools for compressing images and solving inverse problems including denoising, inpainting, and reconstruction from few and noisy measurements. This success can be attributed in part to their ability to represent and generate…
Yuhuai Wu, Yuri Burda, Ruslan Salakhutdinov, Roger Grosse
The past several years have seen remarkable progress in generative models which produce convincing samples of images and other modalities. A shared component of many powerful generative models is a decoder network, a parametric deep neural net that defines a generative distribution. Examples include variational…
Mengran Yan, Chun Tang, Jida Yan, Siti Suhaily Surip + 1 more
Pattern design is essential in various domains, especially in traditional lantern production, where patterns convey cultural history and artistic values. Our research presents an innovative generative model that produces customizable lantern patterns, integrating classical aesthetics with modern design features via a…
Guoqiang Zhong, Wei Gao, Yongbin Liu, Youzhao Yang
—In recent years, research on image generation methods has been developing fast. The auto-encoding variational Bayes method (VAEs) was proposed in 2013, which uses variational inference to learn a latent space from the image database and then generates images using the decoder. The generative adversarial networks…
Xuezhe Ma, Xiang Kong, Shanghang Zhang, Eduard Hovy
In this work, we propose a new generative model that is capable of automatically decoupling global and local representations of images in an entirely unsupervised setting, by embedding a generative flow in the VAE framework to model the decoder. Specifically, the proposed model utilizes the variational auto-encoding…
Aghiles Kebaili, Jérôme Lapuyade-Lahorgue, Su Ruan, Cecilia Di Ruberto + 4 more
Deep learning has become a popular tool for medical image analysis, but the limited availability of training data remains a major challenge, particularly in the medical field where data acquisition can be costly and subject to privacy regulations. Data augmentation techniques offer a solution by artificially increasing…
Minjoo Kim, Yelim Kim, Won Il Park
This study introduces an optical neural network (ONN)-based autoencoder for efficient image processing, utilizing specialized optical matrix-vector multipliers for both encoding and decoding tasks. To address the challenges in efficient decoding, we propose a method that optimizes output processing through scalar…
Aman Singh, Tokunbo Ogunfunmi, Sotiris Kotsiantis
Autoencoders are a self-supervised learning system where, during training, the output is an approximation of the input. Typically, autoencoders have three parts: Encoder (which produces a compressed latent space representation of the input data), the Latent Space (which retains the knowledge in the input data with…
Vignesh Sampath, Iñaki Maurtua, Juan José Aguilar Martín, Aitor Gutierrez
Any computer vision application development starts off by acquiring images and data, then preprocessing and pattern recognition steps to perform a task. When the acquired images are highly imbalanced and not adequate, the desired task may not be achievable. Unfortunately, the occurrence of imbalance problems in…
Tianyang Hu, Fei Chen, Haonan Wang, Jiawei Li + 3 more
'Jiacheng Sun' 'Zhenguo Li'] In generative modeling, numerous successful approaches leverage a lowdimensional latent space, e.g., Stable Diffusion [68] models the latent space induced by an encoder and generates images through a paired decoder. Although the selection of the latent space is empirically pivotal…
Benchen Yang, Xuzhao Liu, Yize Li, Haibo Jin + 2 more
Unpaired image-to-image translation (I2IT) involves establishing an effective mapping between the source and target domains to enable cross-domain image transformation. Previous contrastive learning methods inadequately accounted for the variations in features between two domains and the interrelatedness of elements…
Yuling He, Yingding Zhao, Wenji Yang, Yilu Xu + 1 more
'Juan Pedro Dominguez-Morales'] Due to the sophisticated entanglements for non-rigid deformation, generating person images from source pose to target pose is a challenging work. In this paper, we present a novel framework to generate person images with shape consistency and appearance consistency. The proposed…
Kiran Bacsa, Zhilu Lai, Wei Liu, Michael Todd + 1 more
We propose a new variational autoencoder (VAE) with physical constraints capable of learning the dynamics of Multiple Degree of Freedom (MDOF) dynamic systems. Standard variational autoencoders place greater emphasis on compression than interpretability regarding the learned latent space. We propose a new type of…