Pytorch implementation of Value Iteration Networks (NIPS 2016 best paper)

Last update: Dec 26, 2022

Related tags

Overview

VIN: Value Iteration Networks

A quick thank you

A few others have released amazing related work which helped inspire and improve my own implementation. It goes without saying that this release would not be nearly as good if it were not for all of the following:

Why another VIN implementation?

The Pytorch VIN model in this repository is, in my opinion, more readable and closer to the original Theano implementation than others I have found (both Tensorflow and Pytorch).
This is not simply an implementation of the VIN model in Pytorch, it is also a full Python implementation of the gridworld environments as used in the original MATLAB implementation.
Provide a more extensible research base for others to build off of without needing to jump through the possible MATLAB paywall.

Installation

This repository requires following packages:

SciPy >= 0.19.0
Python >= 2.7 (if using Python 3.x: python3-tk should be installed)
Numpy >= 1.12.1
Matplotlib >= 2.0.0
PyTorch >= 0.1.11

Use pip to install the necessary dependencies:

pip install -U -r requirements.txt

Note that PyTorch cannot be installed directly from PyPI; refer to http://pytorch.org/ for custom installation instructions specific to your needs.

How to train

8x8 gridworld

python train.py --datafile dataset/gridworld_8x8.npz --imsize 8 --lr 0.005 --epochs 30 --k 10 --batch_size 128

16x16 gridworld

python train.py --datafile dataset/gridworld_16x16.npz --imsize 16 --lr 0.002 --epochs 30 --k 20 --batch_size 128

28x28 gridworld

python train.py --datafile dataset/gridworld_28x28.npz --imsize 28 --lr 0.002 --epochs 30 --k 36 --batch_size 128

Flags:

datafile: The path to the data files.
imsize: The size of input images. One of: [8, 16, 28]
lr: Learning rate with RMSProp optimizer. Recommended: [0.01, 0.005, 0.002, 0.001]
epochs: Number of epochs to train. Default: 30
k: Number of Value Iterations. Recommended: [10 for 8x8, 20 for 16x16, 36 for 28x28]
l_i: Number of channels in input layer. Default: 2, i.e. obstacles image and goal image.
l_h: Number of channels in first convolutional layer. Default: 150, described in paper.
l_q: Number of channels in q layer (~actions) in VI-module. Default: 10, described in paper.
batch_size: Batch size. Default: 128

How to test / visualize paths (requires training first)

8x8 gridworld

python test.py --weights trained/vin_8x8.pth --imsize 8 --k 10

16x16 gridworld

python test.py --weights trained/vin_16x16.pth --imsize 16 --k 20

28x28 gridworld

python test.py --weights trained/vin_28x28.pth --imsize 28 --k 36

To visualize the optimal and predicted paths simply pass:

--plot

Flags:

weights: Path to trained weights.
imsize: The size of input images. One of: [8, 16, 28]
plot: If supplied, the optimal and predicted paths will be plotted
k: Number of Value Iterations. Recommended: [10 for 8x8, 20 for 16x16, 36 for 28x28]
l_i: Number of channels in input layer. Default: 2, i.e. obstacles image and goal image.
l_h: Number of channels in first convolutional layer. Default: 150, described in paper.
l_q: Number of channels in q layer (~actions) in VI-module. Default: 10, described in paper.

Results

Gridworld	Sample One	Sample Two
8x8
16x16
28x28

Datasets

Each data sample consists of an obstacle image and a goal image followed by the (x, y) coordinates of current state in the gridworld.

Dataset size	8x8	16x16	28x28
Train set	81337	456309	1529584
Test set	13846	77203	251755

The datasets (8x8, 16x16, and 28x28) included in this repository can be reproduced using the dataset/make_training_data.py script. Note that this script is not optimized and runs rather slowly (also uses a lot of memory :D)

Performance: Success Rate

This is the success rate from rollouts of the learned policy in the environment (taken over 5000 randomly generated domains).

Success Rate	8x8	16x16	28x28
PyTorch	99.69%	96.99%	91.07%

Performance: Test Accuracy

NOTE: This is the accuracy on test set. It is different from the table in the paper, which indicates the success rate from rollouts of the learned policy in the environment.

Test Accuracy	8x8	16x16	28x28
PyTorch	99.83%	94.84%	88.54%

Comments

testing accuracy fairly low

I just tried to follow the instructions in the repo, and tested models trained but got a fairly low accuracy. I'm using pyTorch 0.1.12_1. Is there anything I should pay attention to?

opened by xinleipan 10
Prebuilt Dataset Generation

Hello,

I was wondering how you generated the prebuilt datasets that are downloaded when running download_weights_and_datasets.sh, i.e. what were the max_obs and max_obs_size parameters?

Did you follow this file in the original repo? https://github.com/avivt/VIN/blob/master/scripts/make_data_gridworld_nips.m

Thanks, Emilio

opened by eparisotto 5
the rollout accuracy in test script is lower than the test accuracy in train script.

Hello!

I have a little doubt.Does the rollout accuracy indicate the success rate? If so, why is it lower than the prediction accuracy? In the Aviv's implementation, the success rate of the 8x8 grid world was as high as 99.6%. Why is the success rate in your experiment relatively low?

Thanks!

opened by albzni 4
RUN ERROR

when I run 'python train.py --datafile dataset/gridworld_8x8.npz --imsize 8 --lr 0.005 --epochs 30 --k 10 --batch_size 128', it's ok,but again 'python train.py --datafile dataset/gridworld_16x16.npz --imsize 16 --lr 0.002 --epochs 30 --k 20 --batch_size 128' was run, an error occurred as follows: [email protected]:~/pytorch-value-iteration-networks$ python train.py --datafile dataset/gridworld_16x16.npz --imsize 16 --lr 0.002 --epochs 10 --k 20 --batch_size 128 Traceback (most recent call last): File "train.py", line 135, in config.datafile, imsize=config.imsize, train=True, transform=transform) File "/home/ni/pytorch-value-iteration-networks/dataset/dataset.py", line 22, in init self._process(file, self.train) File "/home/ni/pytorch-value-iteration-networks/dataset/dataset.py", line 58, in _process images = images.astype(np.float32) MemoryError

opened by N-Kingsley 3
Problem of running the test script

Hello,

I downloaded the data with the .sh downloading script you provided, I also got an nps weights file after training. When I ran the testing command I got the following error: Traceback (most recent call last): File "/home/research/DL/VIN/pytorch-value-iteration-networks/test.py", line 158, in main(config) File "/home/research/DL/VIN/pytorch-value-iteration-networks/test.py", line 85, in main _, predictions = vin(X_in, S1_in, S2_in, config) File "/usr/local/lib/python2.7/dist-packages/torch/nn/modules/module.py", line 357, in call result = self.forward(*input, **kwargs) File "/home/research/DL/VIN/pytorch-value-iteration-networks/model.py", line 64, in forward return logits, self.sm(logits) File "/usr/local/lib/python2.7/dist-packages/torch/nn/modules/module.py", line 352, in call for hook in self._forward_pre_hooks.values(): File "/usr/local/lib/python2.7/dist-packages/torch/nn/modules/module.py", line 398, in getattr type(self).name, name)) AttributeError: 'Softmax' object has no attribute '_forward_pre_hooks'

Thanks for helping!

opened by YantianZha 3
Improved readability of the VIN model, in addition to minor changes

My main modification is in the forward method of the model where you extract the q_out from the q values, and not repeating q = F.conv2d(...) in two places. I also made minor improvements, such as adding argparse in the dataset creation script and changing .cuda() into .to(device) in test.py.

opened by shuishida 2

Inconsistent tensor sizes when starting training

Hey there. I'm trying to run

python train.py --datafile dataset/gridworld_8x8.npz --imsize 8 --lr 0.005 --epochs 30 --k 10 --batch_size 128

But I get the following error

Number of Train Samples: 103926
Number of Test Samples: 17434
     Epoch | Train Loss | Train Error | Epoch Time
Traceback (most recent call last):
  File "train.py", line 147, in <module>
    train(net, trainloader, config, criterion, optimizer, use_GPU)
  File "train.py", line 40, in train
    outputs, predictions = net(X, S1, S2, config)
  File "/home/j1k1000o/anaconda3/lib/python3.6/site-packages/torch/nn/modules/module.py", line 224, in __call__
    result = self.forward(*input, **kwargs)
  File "/media/user_home2/j1k1000o/j1k/VINs/pytorch-value-iteration-networks/model.py", line 44, in forward
    q = F.conv2d(torch.cat([r, v], 1), 
  File "/home/j1k1000o/anaconda3/lib/python3.6/site-packages/torch/autograd/variable.py", line 897, in cat
    return Concat.apply(dim, *iterable)
  File "/home/j1k1000o/anaconda3/lib/python3.6/site-packages/torch/autograd/_functions/tensor.py", line 317, in forward
    return torch.cat(inputs, dim)
RuntimeError: inconsistent tensor sizes at /opt/conda/conda-bld/pytorch_1502009910772/work/torch/lib/THC/generic/THCTensorMath.cu:141

I've executed

./download_weights_and_datasets.sh

as well as

python ./dataset/make_training_data.py

And I'm running it on an Ubuntu 16.04, python 3.6 and with all the requirements installed.

Can you help me out?

opened by juancprzs 2

Don't understand VIN last step

    slice_s1 = S1.long().expand(config.imsize, 1, config.l_q, q.size(0))
    slice_s1 = slice_s1.permute(3, 2, 1, 0)
    q_out = q.gather(2, slice_s1).squeeze(2)

What does this 3 lines do?

opened by QiXuanWang 1

KeyError: 'arr_1 is not a file in the archive'

python3 train.py --datafile dataset/gridworld_8x8.npz --imsize 8 --lr 0.005 --epochs 30 --k 10 --batch_size 128 Traceback (most recent call last): File "train.py", line 135, in config.datafile, imsize=config.imsize, train=True, transform=transform) File "/home/user/pytorch/tutorials/valueiterationnetworks/pytorch-value-iteration-networks/dataset/dataset.py", line 22, in init self._process(file, self.train) File "/home/user/pytorch/tutorials/valueiterationnetworks/pytorch-value-iteration-networks/dataset/dataset.py", line 49, in _process S1 = f['arr_1'] File "/home/user/miniconda3/lib/python3.6/site-packages/numpy/lib/npyio.py", line 255, in getitem raise KeyError("%s is not a file in the archive" % key) KeyError: 'arr_1 is not a file in the archive'

I got this error, could you please

opened by derelearnro 1
Problem of running dataset/make_training_data.py script
Hi

When I tried to run the make_training_data.py script to generate the gridworld.npz file, I got the following error:

FileNotFoundError: [Errno 2] No such file or directory: 'dataset/gridworld_28x28.npz'

And I found that line 101 should be modified as follows:

save_path = "gridworld_{0}x{1}".format(dom_size[0], dom_size[1])
opened by ruqing00 0

Releases(v1.1)

v1.1(Apr 18, 2018)
Patches for newest PyTorch releases.

Fixes #3 #4 #5

Source code(tar.gz)
Source code(zip)
gridworld_16x16.npz(10.21 MB)
gridworld_28x28.npz(253.55 MB)
gridworld_8x8.npz(566.12 KB)
vin_16x16.pth(13.54 KB)
vin_28x28.pth(13.54 KB)
vin_8x8.pth(13.54 KB)
v1.0(Apr 21, 2017)

This release includes pre-built datasets and pre-trained models for the 8x8, 16x16, and 28x28 gridworlds.
Source code(tar.gz)
Source code(zip)
gridworld_16x16.npz(33.58 MB)
gridworld_28x28.npz(253.55 MB)
gridworld_8x8.npz(1.28 MB)
vin_16x16.pth(24.37 KB)
vin_28x28.pth(24.42 KB)
vin_8x8.pth(24.37 KB)

Owner

Kent Sommer

Software Engineer @ Toyota Research Institute (SF Bay Area)

GitHub Repository

Face-Recognition-Attendence-System - This face recognition Attendence system using Python

Face-Recognition-Attendence-System I have developed this face recognition Attend

4 May 10, 2022

Re-implementation of the vector capsule with dynamic routing

VectorCapsule Re-implementation of the vector capsule with dynamic routing We implement the vector capsule and dynamic routing via graph neural networ

10 Feb 10, 2022

Self-Correcting Quantum Many-Body Control using Reinforcement Learning with Tensor Networks

Self-Correcting Quantum Many-Body Control using Reinforcement Learning with Tensor Networks This repository contains the code and data for the corresp

7 Apr 23, 2022

This is the unofficial code of Deep Dual-resolution Networks for Real-time and Accurate Semantic Segmentation of Road Scenes. which achieve state-of-the-art trade-off between accuracy and speed on cityscapes and camvid, without using inference acceleration and extra data

Deep Dual-resolution Networks for Real-time and Accurate Semantic Segmentation of Road Scenes Introduction This is the unofficial code of Deep Dual-re

113 Dec 23, 2022

Two-stage CenterNet

Probabilistic two-stage detection Two-stage object detectors that use class-agnostic one-stage detectors as the proposal network. Probabilistic two-st

1.1k Jan 03, 2023

official code for dynamic convolution decomposition

Revisiting Dynamic Convolution via Matrix Decomposition (ICLR 2021) A pytorch implementation of DCD. If you use this code in your research please cons

110 Nov 23, 2022

This repo contains source code and materials for the TEmporally COherent GAN SIGGRAPH project.

TecoGAN This repository contains source code and materials for the TecoGAN project, i.e. code for a TEmporally COherent GAN for video super-resolution

5.2k Jan 02, 2023

A Large Scale Benchmark for Individual Treatment Effect Prediction and Uplift Modeling

large-scale-ITE-UM-benchmark This repository contains code and data to reproduce the results of the paper "A Large Scale Benchmark for Individual Trea

10 Nov 19, 2022

JAX code for the paper "Control-Oriented Model-Based Reinforcement Learning with Implicit Differentiation"

Optimal Model Design for Reinforcement Learning This repository contains JAX code for the paper Control-Oriented Model-Based Reinforcement Learning wi

43 Sep 28, 2022

A full-fledged version of Pix2Seq

Stable-Pix2Seq A full-fledged version of Pix2Seq What it is. This is a full-fledged version of Pix2Seq. Compared with unofficial-pix2seq, stable-pix2s

205 Dec 27, 2022

Repository for the Bias Benchmark for QA dataset.

BBQ Repository for the Bias Benchmark for QA dataset. Authors: Alicia Parrish, Angelica Chen, Nikita Nangia, Vishakh Padmakumar, Jason Phang, Jana Tho

18 Nov 18, 2022

✨风纪委员会自动投票脚本，利用Github Action帮你进行裁决操作（为了让其他风纪委员有案件可判，本程序从中午12点才开始运行，有需要请自己修改运行时间）

风纪委员会自动投票本脚本通过使用Github Action来实现B站风纪委员的自动投票功能，喜欢请给我点个STAR吧！如果你不是风纪委员，在符合风纪委员申请条件的情况下，本脚本会自动帮你申请投票时间是早上八点，如果有需要请自行修改.github/workflows/Judge.yml中的时间，

25 Feb 17, 2021

[ICCV'21] Official implementation for the paper Social NCE: Contrastive Learning of Socially-aware Motion Representations

CrowdNav with Social-NCE This is an official implementation for the paper Social NCE: Contrastive Learning of Socially-aware Motion Representations by

125 Dec 23, 2022

Large scale and asynchronous Hyperparameter Optimization at your fingertip.

Syne Tune This package provides state-of-the-art distributed hyperparameter optimizers (HPO) where trials can be evaluated with several backend option

236 Jan 01, 2023

Self-driving car env with PPO algorithm from stable baseline3

Self-driving car with RL stable baseline3 Most of the project develop from https://github.com/GerardMaggiolino/Gym-Medium-Post Please check it out! Th

7 Dec 22, 2022

In this tutorial, you will perform inference across 10 well-known pre-trained object detectors and fine-tune on a custom dataset. Design and train your own object detector.

Object Detection Object detection is a computer vision task for locating instances of predefined objects in images or videos. In this tutorial, you wi

62 Dec 25, 2022

This project provides an unsupervised framework for mining and tagging quality phrases on text corpora with pretrained language models (KDD'21).

UCPhrase: Unsupervised Context-aware Quality Phrase Tagging To appear on KDD'21...[pdf] This project provides an unsupervised framework for mining and

146 Dec 22, 2022

Code for CoMatch: Semi-supervised Learning with Contrastive Graph Regularization

CoMatch: Semi-supervised Learning with Contrastive Graph Regularization (Salesforce Research) This is a PyTorch implementation of the CoMatch paper [B

107 Dec 14, 2022

This repository contains the implementation of the paper Contrastive Instance Association for 4D Panoptic Segmentation using Sequences of 3D LiDAR Scans

Contrastive Instance Association for 4D Panoptic Segmentation using Sequences of 3D LiDAR Scans This repository contains the implementation of the pap

40 Dec 01, 2022

Stochastic gradient descent with model building

Stochastic Model Building (SMB) This repository includes a new fast and robust stochastic optimization algorithm for training deep learning models. Th

22 Jan 19, 2022

Pytorch implementation of Value Iteration Networks (NIPS 2016 best paper)

Related tags

Overview

VIN: Value Iteration Networks

A quick thank you

Why another VIN implementation?

Installation

How to train

8x8 gridworld

16x16 gridworld

28x28 gridworld

How to test / visualize paths (requires training first)

8x8 gridworld

16x16 gridworld

28x28 gridworld

Results

Datasets

Performance: Success Rate

Performance: Test Accuracy

Comments

Releases(v1.1)

v1.1(Apr 18, 2018)

v1.0(Apr 21, 2017)

Owner

Kent Sommer

Face-Recognition-Attendence-System - This face recognition Attendence system using Python

Re-implementation of the vector capsule with dynamic routing

Self-Correcting Quantum Many-Body Control using Reinforcement Learning with Tensor Networks

This is the unofficial code of Deep Dual-resolution Networks for Real-time and Accurate Semantic Segmentation of Road Scenes. which achieve state-of-the-art trade-off between accuracy and speed on cityscapes and camvid, without using inference acceleration and extra data

Two-stage CenterNet

official code for dynamic convolution decomposition

This repo contains source code and materials for the TEmporally COherent GAN SIGGRAPH project.

A Large Scale Benchmark for Individual Treatment Effect Prediction and Uplift Modeling

JAX code for the paper "Control-Oriented Model-Based Reinforcement Learning with Implicit Differentiation"

A full-fledged version of Pix2Seq

Repository for the Bias Benchmark for QA dataset.

✨风纪委员会自动投票脚本，利用Github Action帮你进行裁决操作（为了让其他风纪委员有案件可判，本程序从中午12点才开始运行，有需要请自己修改运行时间）

[ICCV'21] Official implementation for the paper Social NCE: Contrastive Learning of Socially-aware Motion Representations

Large scale and asynchronous Hyperparameter Optimization at your fingertip.

Self-driving car env with PPO algorithm from stable baseline3

In this tutorial, you will perform inference across 10 well-known pre-trained object detectors and fine-tune on a custom dataset. Design and train your own object detector.

This project provides an unsupervised framework for mining and tagging quality phrases on text corpora with pretrained language models (KDD'21).

Code for CoMatch: Semi-supervised Learning with Contrastive Graph Regularization

This repository contains the implementation of the paper Contrastive Instance Association for 4D Panoptic Segmentation using Sequences of 3D LiDAR Scans

Stochastic gradient descent with model building