This repository contains the data and code for the paper "Diverse Text Generation via Variational Encoder-Decoder Models with Gaussian Process Priors" ([email protected])

Last update: Dec 29, 2022

Overview

GP-VAE

This repository provides datasets and code for preprocessing, training and testing models for the paper:

Diverse Text Generation via Variational Encoder-Decoder Models with Gaussian Process Priors
Wanyu Du, Jianqiao Zhao, Liwei Wang and Yangfeng Ji
ACL 2022 6th Workshop on Structured Prediction for NLP

Installation

The following command installs all necessary packages:

pip install -r requirements.txt

The project was tested using Python 3.6.6.

Datasets

Twitter URL includes trn/val/tst.tsv, which has the following format in each line:

source_sentence \t reference_sentence

GYAFC has two sub-domains em and fr, please request and download the data from the original paper here.

Models

Training

Train the LSTM-based variational encoder-decoder with GP priors:

cd models/pg/
python main.py --task train --data_file ../../data/twitter_url \
			   --model_type gp_full --kernel_v 65.0 --kernel_r 0.0001

where --data_file indicates the data path for the training data,
--model_type indicates which prior to use, including copynet/normal/gp_full,
--kernel_v and --kernel_r specifies the hyper-parameters for the kernel of GP prior.

Train the transformer-based variational encoder-decoder with GP priors:

cd models/t5/
python t5_gpvae.py --task train --dataset twitter_url \
    			   --kernel_v 512.0 --kernel_r 0.001

where --data_file indicates the data path for the training data,
--kernel_v and --kernel_r specifies the hyper-parameters for the kernel of GP prior.

Inference

Test the LSTM-based variational encoder-decoder with GP priors:

cd models/pg/
python main.py --task decode --data_file ../../data/twitter_url \
			   --model_type gp_full --kernel_v 65.0 --kernel_r 0.0001 \
			   --decode_from sample \
			   --model_file /path/to/best/checkpoint

where --data_file indicates the data path for the testing data,
--model_type indicates which prior to use, including copynet/normal/gp_full,
--kernel_v and --kernel_r specifies the hyper-parameters for the kernel of GP prior,
--decode_from indicates generating results conditioning on z_mean or randomly sampled z, including mean/sample.

Test the transformer-based variational encoder-decoder with GP priors:

cd models/t5/
python t5_gpvae.py --task eval --dataset twitter_url \
    			   --kernel_v 512.0 --kernel_r 0.001 \
    			   --from_mean \
    			   --timestamp '2021-02-14-04-57-04' \
    			   --ckpt '30000' # load best checkpoint

where --data_file indicates the data path for the testing data,
--kernel_v and --kernel_r specifies the hyper-parameters for the kernel of GP prior,
--from_mean indicates whether to generate results conditioning on z_mean or randomly sampled z,
--timestamp and --ckpt indicate the file path for the best checkpoint.

Citation

If you find this work useful for your research, please cite our paper:

Diverse Text Generation via Variational Encoder-Decoder Models with Gaussian Process Priors

@inproceedings{du2022gpvae,
    title = "Diverse Text Generation via Variational Encoder-Decoder Models with Gaussian Process Priors",
    author = "Du, Wanyu and Zhao, Jianqiao and Wang, Liwei and Ji, Yangfeng",
    booktitle = "Proceedings of the 6th Workshop on Structured Prediction for NLP (SPNLP 2022)",
    year = "2022",
    publisher = "Association for Computational Linguistics",
}

This repository contains the data and code for the paper "Diverse Text Generation via Variational Encoder-Decoder Models with Gaussian Process Priors" ([email protected])

Related tags

Overview

GP-VAE

Installation

Datasets

Models

Training

Inference

Citation

Diverse Text Generation via Variational Encoder-Decoder Models with Gaussian Process Priors

Owner

Wanyu Du

KGDet: Keypoint-Guided Fashion Detection (AAAI 2021)

This is an official implementation for "Video Swin Transformers".

A modern pure-Python library for reading PDF files

Official code repository for ICCV 2021 paper: Gravity-Aware Monocular 3D Human Object Reconstruction

Illuminated3D This project participates in the Nasa Space Apps Challenge 2021.

Self-describing JSON-RPC services made easy

[3DV 2021] Channel-Wise Attention-Based Network for Self-Supervised Monocular Depth Estimation

A PaddlePaddle implementation of Time Interval Aware Self-Attentive Sequential Recommendation.

Reimplementation of the paper "Attention, Learn to Solve Routing Problems!" in jax/flax.

A forwarding MPI implementation that can use any other MPI implementation via an MPI ABI

Tracking code for the winner of track 1 in the MMP-Tracking Challenge at ICCV 2021 Workshop.

Shape Matching of Real 3D Object Data to Synthetic 3D CADs (3DV project @ ETHZ)

ML course - EPFL Machine Learning Course, Fall 2021

Deep Learning & 3D Convolutional Neural Networks for Speaker Verification

Annotate datasets with a semi-trained or fully trained YOLOv5 model

Scripts and a shader to get you started on setting up an exported Koikatsu character in Blender.

A pytorch &keras implementation and demo of Fastformer.

Generalized hybrid model for mode-locked laser diodes with an extended passive cavity

GMFlow: Learning Optical Flow via Global Matching

A hyperparameter optimization framework