RL-GAN: Transfer Learning for Related Reinforcement Learning Tasks via Image-to-Image Translation

Last update: Nov 10, 2022

Related tags

Overview

RL-GAN: Transfer Learning for Related Reinforcement Learning Tasks via Image-to-Image Translation

RL-GAN is an official implementation of the paper: Transfer Learning for Related Reinforcement Learning Tasks via Image-to-Image Translation.

Paper

Shani Gamrian, Yoav Goldberg, "Transfer Learning for Related Reinforcement Learning Tasks via Image-to-Image Translation"

@article{DBLP:journals/corr/abs-1806-07377,
  author    = {Shani Gamrian and
               Yoav Goldberg},
  title     = {Transfer Learning for Related Reinforcement Learning Tasks via Image-to-Image
               Translation},
  journal   = {CoRR},
  volume    = {abs/1806.07377},
  year      = {2018},
  url       = {http://arxiv.org/abs/1806.07377},
  archivePrefix = {arXiv},
  eprint    = {1806.07377},
  timestamp = {Mon, 13 Aug 2018 16:48:23 +0200},
  biburl    = {https://dblp.org/rec/bib/journals/corr/abs-1806-07377},
  bibsource = {dblp computer science bibliography, https://dblp.org}
}

Videos:

Breakout

RoadFighter

Installation

The code was tested on Ubuntu 16.04 with Python 3.6
Install packages by typing the command: pip install -r requirements.txt.
For Road Fighter, clone and install the repo: https://github.com/ShaniGam/retro

Getting Started

Breakout Examples

Train Breakout from scratch:

python -m breakout_a3c.main --num-processes 32 --variation 'standart'

Transfer from standart to diagonals variation and fine-tune the model:

python -m breakout_a3c.main --num-processes 32 --variation diagonals --ft-setting full-ft --test

Collect images for UNIT training:

python -m breakout_a3c.main --collect-images --num-collected-imgs 100000 --variation diagonals --num-processes 1

Train UNIT:

python -m unit.train --trainer UNIT --config unit/configs/breakout-diagonals.yaml

Run Breakout with UNIT:

python -m breakout_a3c.main --variation diagonals --test --ft-setting full-ft --test-gan --gan-dir breakout-diagonals --num-processes 0

Road Fighter Examples

Train level 1 of Road Fighter

python -m roadfighter_a2c.main --num-processes 84

Collect images for UNIT training:

python -m roadfighter_a2c.main -level 1 --collect-images --num-collected-imgs 100000 --num-processes 1
python -m roadfighter_a2c.main -level 2 --collect-images --num-collected-imgs 100000 --num-processes 1

Train UNIT:

python -m unit.train --trainer UNIT --config unit/configs/roadfighter-lvl2.yaml

Run Road Fighter with UNIT:

python -m roadfighter_a2c.main --load --level 2 --test-gan --gan-dir roadfighter-lvl2-kl01 --num-processes 1

Run Road Fighter with UNIT and Imitation Learning:

python -m roadfighter_a2c.main_imitation --load --gan-dir roadfighter-lvl2-kl01 --gan-imitation-file '00320000' --log-name lvl2.log --super-during-rl --level 2 --det-score 5350

Acknowledgments

The code was written by Shani Gamrian and is based on the repositories: pytorch-a3c, pytorch-a2c, UNIT

TO-DO

Add links for pretrained models.
Create videos.

RL-GAN: Transfer Learning for Related Reinforcement Learning Tasks via Image-to-Image Translation

Related tags

Overview

RL-GAN: Transfer Learning for Related Reinforcement Learning Tasks via Image-to-Image Translation

Paper

Videos:

Installation

Getting Started

Breakout Examples

Road Fighter Examples

Acknowledgments

TO-DO

Owner

Dataset for the Research2Clinics @ NeurIPS 2021 Paper: What Do You See in this Patient? Behavioral Testing of Clinical NLP Models

face property detection pytorch

ExCon: Explanation-driven Supervised Contrastive Learning

Utilities and information for the signals.numer.ai tournament

YOLOV4运行在嵌入式设备上

A platform for intelligent agent learning based on a 3D open-world FPS game developed by Inspir.AI.

Image classification for projects and researches

Mosaic of Object-centric Images as Scene-centric Images (MosaicOS) for long-tailed object detection and instance segmentation.

1st-in-MICCAI2020-CPM - Combined Radiology and Pathology Classification

Pytorch implementation of FlowNet by Dosovitskiy et al.

DABO: Data Augmentation with Bilevel Optimization

Modifications of the official PyTorch implementation of StyleGAN3. Let's easily generate images and videos with StyleGAN2/2-ADA/3!

Implementation of Lie Transformer, Equivariant Self-Attention, in Pytorch

A curated list of references for MLOps

计算机视觉中用到的注意力模块和其他即插即用模块PyTorch Implementation Collection of Attention Module and Plug&Play Module

Object tracking and object detection is applied to track golf puts in real time and display stats/games.

SpiroMask: Measuring Lung Function Using Consumer-Grade Masks

Python script to download the celebA-HQ dataset from google drive

Image-Stitching - Panorama composition using SIFT Features and a custom implementaion of RANSAC algorithm

The code for Expectation-Maximization Attention Networks for Semantic Segmentation (ICCV'2019 Oral)