Pytorch implementation of VAEs for heterogeneous likelihoods.

Last update: Nov 29, 2022

Related tags

Overview

Heterogeneous VAEs

Beware: This repository is under construction 🛠️

Pytorch implementation of different VAE models to model heterogeneous data. Here, we call heterogeneous data those for which we assume that each feature is of a different type, and therefore each feature is assumed to have a different likelihood. Heterogeneous data is also known as mixed-type data and tabular data.

Usage

This repository is not meant to be a library which you can install and use as it is, but rather as a ML project code which you can freely fork and modify to fit your particular needs.

Dependencies

We are working on providing a conda requirements file. For the moment, there is a Dockerfile which you can build and use, or simply look at the project dependencies from there.

Example

You can find information about all the available arguments via python main.py --help. For example, you can train the Wine dataset on a heterogeneous VAE with default arguments using:

python main.py -model=vae -dataset=datasets/Wine -seed=2 -miss-perc=20 -miss-suffix=1

Models

This repository contains implementations of the following models, adapted for heterogeneous likelihoods (if you use them in your work, make sure to cite the original authors):

Autoencoding Variational Bayes (VAE): https://arxiv.org/abs/1312.6114
Importance Weighted Autoencoder (IWAE): http://arxiv.org/abs/1509.00519
Doubly Reparametrized Gradient Estimators for Monte Carlo objectives (DReG): http://arxiv.org/abs/1810.04152
Handling Incomplete Heterogeneous Data using VAEs (HI-VAE): http://arxiv.org/abs/1807.03653

Likelihoods

The code supports the following likelihoods at the moment:

Gaussian, for real-valued features.
Log-normal, for positive-valued real features.
Bernoulli, for binary features.
Categorical, for categorical features.
Poisson, for positive-valued integer (count) features.

Datasets

We provide with this code some example datasets taken from UCI and R package datasets. You can use any dataset as long as the format is the same.

Contributing

The code can be further simplified and polished, and we still have some legacy code. Pull requests and issues are more than welcome, as long as it contributes to making the code clean, simple, general, and elegant.

Pytorch implementation of VAEs for heterogeneous likelihoods.

Related tags

Overview

Heterogeneous VAEs

Usage

Dependencies

Example

Models

Likelihoods

Datasets

Contributing

Owner

Adrián Javaloy

Benchmark for the generalization of 3D machine learning models across different remeshing/samplings of a surface.

MEDS: Enhancing Memory Error Detection for Large-Scale Applications

PyDeepFakeDet is an integrated and scalable tool for Deepfake detection.

Neural network for digit classification powered by cuda

Exploit Camera Raw Data for Video Super-Resolution via Hidden Markov Model Inference

Adversarial Learning for Modeling Human Motion

Cognition-aware Cognate Detection

Code used to generate the results appearing in "Train longer, generalize better: closing the generalization gap in large batch training of neural networks"

DrWhy is the collection of tools for eXplainable AI (XAI). It's based on shared principles and simple grammar for exploration, explanation and visualisation of predictive models.

Generic ecosystem for feature extraction from aerial and satellite imagery

Official implementation of Pixel-Level Bijective Matching for Video Object Segmentation

naked is a Python tool which allows you to strip a model and only keep what matters for making predictions.

Official code of paper "PGT: A Progressive Method for Training Models on Long Videos" on CVPR2021

CDTrans: Cross-domain Transformer for Unsupervised Domain Adaptation

A custom-designed Spider Robot trained to walk using Deep RL in a PyBullet Simulation

TorchOk - The toolkit for fast Deep Learning experiments in Computer Vision

Styled text-to-drawing synthesis method. Featured at the 2021 NeurIPS Workshop on Machine Learning for Creativity and Design

Airborne Optical Sectioning (AOS) is a wide synthetic-aperture imaging technique

PyTorch implementation for "HyperSPNs: Compact and Expressive Probabilistic Circuits", NeurIPS 2021

Godot RL Agents is a fully Open Source packages that allows video game creators