Unofficial JAX implementations of Deep Learning models

Last update: Jan 05, 2023

Overview

JAX Models

Table of Contents

About The Project
Getting Started
Contributing
License
Contact

About The Project

The JAX Models repository aims to provide open sourced JAX/Flax implementations for research papers originally without code or code written with frameworks other than JAX. The goal of this project is to make a collection of models, layers, activations and other utilities that are most commonly used for research. All papers and derived or translated code is cited in either the README or the docstrings. If you think that any citation is missed then please raise an issue.

All implementations provided here are available on Papers With Code.

Available model implementations for JAX are:

MetaFormer is Actually What You Need for Vision (Weihao Yu et al., 2021)
Augmenting Convolutional networks with attention-based aggregation (Hugo Touvron et al., 2021)
MPViT : Multi-Path Vision Transformer for Dense Prediction (Youngwan Lee et al., 2021)
MLP-Mixer: An all-MLP Architecture for Vision (Ilya Tolstikhin et al., 2021)
Patches Are All You Need (Anonymous et al., 2021)
SegFormer: Simple and Efficient Design for Semantic Segmentation with Transformers (Enze Xie et al., 2021)
A ConvNet for the 2020s (Zhuang Liu et al., 2021)
Masked Autoencoders Are Scalable Vision Learners (Kaiming He et al., 2021)

Available layers for out-of-the-box integration:

DropPath (Stochastic Depth) (Gao Huang et al., 2021)
Squeeze-and-Excitation Layer (Jie Hu et al. 2019)
Depthwise Convolution (François Chollet, 2017)

Prerequisites

Prerequisites can be installed separately through the requirements.txt file in the main directory using:

pip install -r requirements.txt

The use of a virtual environment is highly recommended to avoid version incompatibilites.

Installation

This project is built with Python 3 for the latest JAX/Flax versions and can be directly installed via pip.

pip install jax-models

If you wish to use the latest version then you can directly clone the repository too.

git clone https://github.com/DarshanDeshpande/jax-models.git

Usage

To see all model architectures available:

from jax_models.models.model_registry import list_models
from pprint import pprint

pprint(list_models())

To load your desired model:

from jax_models.models.model_registry import load_model
load_model('mpvit-base', attach_head=True, num_classes=1000, dropout=0.1)

Contributing

Please raise an issue if any implementation gives incorrect results, crashes unexpectedly during training/inference or if any citation is missing.

You can contribute to jax_models by supporting me with compute resources or by contributing your own resources to provide pretrained weights.

If you wish to donate to this inititative then please drop me a mail here.

License

Distributed under the Apache 2.0 License. See LICENSE for more information.

Contact

Feel free to reach out for any issues or requests related to these implementations

Darshan Deshpande - Email | Twitter | LinkedIn

You might also like...

Very deep VAEs in JAX/Flax

Very Deep VAEs in JAX/Flax Implementation of the experiments in the paper Very Deep VAEs Generalize Autoregressive Models and Can Outperform Them on I

42 Dec 12, 2022

Conservative Q Learning for Offline Reinforcement Reinforcement Learning in JAX

CQL-JAX This repository implements Conservative Q Learning for Offline Reinforcement Reinforcement Learning in JAX (FLAX). Implementation is built on

8 Nov 7, 2022

PyTorch implementations of neural network models for keyword spotting

Honk: CNNs for Keyword Spotting Honk is a PyTorch reimplementation of Google's TensorFlow convolutional neural networks for keyword spotting, which ac

475 Dec 15, 2022

Unofficial implementation of Proxy Anchor Loss for Deep Metric Learning

Proxy Anchor Loss for Deep Metric Learning Unofficial pytorch, tensorflow and mxnet implementations of Proxy Anchor Loss for Deep Metric Learning. Not

3 Jun 9, 2021

Time-series-deep-learning - Developing Deep learning LSTM, BiLSTM models, and NeuralProphet for multi-step time-series forecasting of stock price.

Stock Price Prediction Using Deep Learning Univariate Time Series Predicting stock price using historical data of a company using Neural networks for

7 Nov 27, 2022

FedJAX is a library for developing custom Federated Learning (FL) algorithms in JAX.

FedJAX: Federated learning with JAX What is FedJAX? FedJAX is a library for developing custom Federated Learning (FL) algorithms in JAX. FedJAX priori

208 Dec 14, 2022

Objax Apache-2Objax (🥉19 · ⭐ 580) - Objax is a machine learning framework that provides an Object.. Apache-2 jax

Objax Tutorials | Install | Documentation | Philosophy This is not an officially supported Google product. Objax is an open source machine learning fr

729 Jan 2, 2023

Plug-n-Play Reinforcement Learning in Python with OpenAI Gym and JAX

coax is built on top of JAX, but it doesn't have an explicit dependence on the jax python package. The reason is that your version of jaxlib will depend on your CUDA version.

128 Dec 27, 2022

JAX code for the paper "Control-Oriented Model-Based Reinforcement Learning with Implicit Differentiation"

Optimal Model Design for Reinforcement Learning This repository contains JAX code for the paper Control-Oriented Model-Based Reinforcement Learning wi

43 Sep 28, 2022

Comments

Missing Axis Swap in ExtractPatches and MergePatches

In patch_utils.py, the modules ExtractPatches and MergePatches are missing an axis swap between the reshapes, resulting in the extracted patches becoming horizontal stripes. For example, if we follow the code in ExtractPatches:

>>> inputs = jnp.arange(16).reshape(1, 4, 4, 1)
>>> inputs[0, :, :, 0]

DeviceArray([[ 0,  1,  2,  3],
             [ 4,  5,  6,  7],
             [ 8,  9, 10, 11],
             [12, 13, 14, 15]], dtype=int32)

>>> patch_size = 2
>>> batch, height, width, channels = inputs.shape
>>> height, width = height // patch_size, width // patch_size
>>> x = jnp.reshape(inputs, (batch, height, patch_size, width, patch_size, channels))
>>> x = jnp.reshape(x, (batch, height * width, patch_size ** 2 * channels))
>>> x[0, 0, :]

DeviceArray([0, 1, 2, 3], dtype=int32)

We see that the first patch extracted is not the patch containing [0, 1, 4, 5], but the horizontal stripe [0, 1, 2, 3]. To fix this problem, we should add an axis swap. For ExtractPatches, this should be:

batch, height, width, channels = inputs.shape
height, width = height // patch_size, width // patch_size
x = jnp.reshape(
    inputs, (batch, height, patch_size, width, patch_size, channels)
)
x = jnp.swapaxes(x, 2, 3)
x = jnp.reshape(x, (batch, height * width, patch_size ** 2 * channels))

For MergePatches, this should be:

batch, length, _ = inputs.shape
height = width = int(length**0.5)
x = jnp.reshape(inputs, (batch, height, width, patch_size, patch_size, -1))
x = jnp.swapaxes(x, 2, 3)
x = jnp.reshape(x, (batch, height * patch_size, width * patch_size, -1))

bug

opened by young-geng 4

fix convnext to make it work with jax.jit

Hey, first of all, thanks for the nice codebase. When doing inference using the convnext model, I noticed the following issue:

Calling x.item() will call float(x), which breaks the jit tracer. We can remove the list comprehension in unnecessary conversion to make jax.jit work. Without jax.jit, the model is very slow for me, running with only ~30% GPU utilization (RTX 3090).

This issue could apply to other models as well, maybe it is a good idea to include a test for applying jax.jit to each model?

opened by maxidl 1

Releases(v0.5-van)

v0.5-van(Feb 27, 2022)

Weights for Visual Attention Network (Meng-Hao Guo et al., 2022). All weights translated from the official repository. Full credits go to the original authors.
Source code(tar.gz)
Source code(zip)
van_base.weights(101.50 MB)
van_large.weights(170.97 MB)
van_small.weights(52.93 MB)
van_tiny.weights(15.70 MB)
v0.4-cait(Feb 18, 2022)

Weights for Going deeper with Image Transformers (Hugo Touvron et al., 2021) These weights have been translated from the official Github repository and all credits for the weights go to the original authors.
Source code(tar.gz)
Source code(zip)
cait_m36_384.weights(1034.64 MB)
cait_m48_448.weights(1359.81 MB)
cait_s24_224.weights(178.98 MB)
cait_s24_384.weights(179.54 MB)
cait_s36_384.weights(260.81 MB)
cait_xs24_384.weights(101.75 MB)
cait_xxs24_224.weights(45.62 MB)
cait_xxs24_384.weights(45.90 MB)
cait_xxs36_224.weights(66.01 MB)
cait_xxs36_384.weights(66.29 MB)
v0.3-pvit(Feb 11, 2022)

Weights from Pyramid Vision Transformer: A Versatile Backbone for Dense Prediction without Convolutions (Wenhai Wang et al., 2021). All credits for these weights go to the original authors.
Source code(tar.gz)
Source code(zip)
pvit_b0.weights(13.99 MB)
pvit_b1.weights(53.44 MB)
pvit_b2.weights(96.76 MB)
pvit_b2_linear.weights(86.04 MB)
pvit_b3.weights(172.58 MB)
pvit_b4.weights(238.65 MB)
pvit_b5.weights(312.66 MB)
v0.2-convnext(Feb 6, 2022)

Weights for ConvNeXt (Zhuang Liu et al, 2022) translated from the official repository.

All credits for the weights go to the original authors.
Source code(tar.gz)
Source code(zip)
convnext_base_224_1k.weights(337.96 MB)
convnext_base_224_22k.weights(419.45 MB)
convnext_base_224_22k_1k.weights(337.96 MB)
convnext_base_384_1k.weights(337.96 MB)
convnext_base_384_22k_1k.weights(337.96 MB)
convnext_large_224_1k.weights(754.43 MB)
convnext_large_224_22k.weights(876.62 MB)
convnext_large_224_22k_1k.weights(754.43 MB)
convnext_large_384_1k.weights(754.43 MB)
convnext_large_384_22k_1k.weights(754.43 MB)
convnext_small_224_1k.weights(191.59 MB)
convnext_tiny_224_1k.weights(109.06 MB)
convnext_xlarge_224_22k.weights(1498.80 MB)
convnext_xlarge_224_22k_1k.weights(1335.90 MB)
convnext_xlarge_384_22k_1k.weights(1335.90 MB)
v0.1-swin(Jan 24, 2022)

This release contains weights for the entire stack of Swin Transformer models (SwinTiny224, SwinSmall224, SwinBase224, SwinBase384, SwinLarge224, SwinLarge384).

These weights have been ported from the official repository and timm.
Source code(tar.gz)
Source code(zip)
swin_base_224_22k.weights(416.30 MB)
swin_base_384_22k.weights(416.82 MB)
swin_large_224_22k.weights(871.91 MB)
swin_large_384_22k.weights(872.69 MB)
swin_small_224_1k.weights(189.24 MB)
swin_tiny_224_1k.weights(107.91 MB)

Owner

Helping Machines Learn Better 💻😃

GitHub Repository

The first dataset of composite images with rationality score indicating whether the object placement in a composite image is reasonable.

Object-Placement-Assessment-Dataset-OPA Object-Placement-Assessment (OPA) is to verify whether a composite image is plausible in terms of the object p

53 Nov 15, 2022

Neural Network to colorize grayscale images

#colornet Neural Network to colorize grayscale images Results Grayscale Prediction Ground Truth Eiji K used colornet for anime colorization Sources Au

3.6k Dec 24, 2022

Improving Contrastive Learning by Visualizing Feature Transformation, ICCV 2021 Oral

Improving Contrastive Learning by Visualizing Feature Transformation This project hosts the codes, models and visualization tools for the paper: Impro

83 Dec 15, 2022

Repository containing detailed experiments related to the paper "Memotion Analysis through the Lens of Joint Embedding".

Memotion Analysis Through The Lens Of Joint Embedding This repository contains the experiments conducted as described in the paper 'Memotion Analysis

1 Mar 16, 2022

Code for SALT: Stackelberg Adversarial Regularization, EMNLP 2021.

SALT: Stackelberg Adversarial Regularization Code for Adversarial Regularization as Stackelberg Game: An Unrolled Optimization Approach, EMNLP 2021. R

10 Jan 10, 2022

Invariant Causal Prediction for Block MDPs

MISA Abstract Generalization across environments is critical to the successful application of reinforcement learning algorithms to real-world challeng

41 Sep 17, 2022

FastyAPI is a Stack boilerplate optimised for heavy loads.

FastyAPI A FastAPI based Stack boilerplate for heavy loads. Explore the docs » View Demo · Report Bug · Request Feature Table of Contents About The Pr

47 Dec 27, 2022

Classification models 1D Zoo - Keras and TF.Keras

Classification models 1D Zoo - Keras and TF.Keras This repository contains 1D variants of popular CNN models for classification like ResNets, DenseNet

12 Jan 06, 2023

Scalable, Portable and Distributed Gradient Boosting (GBDT, GBRT or GBM) Library, for Python, R, Java, Scala, C++ and more. Runs on single machine, Hadoop, Spark, Dask, Flink and DataFlow

eXtreme Gradient Boosting Community | Documentation | Resources | Contributors | Release Notes XGBoost is an optimized distributed gradient boosting l

23.6k Dec 31, 2022

Unofficial JAX implementations of Deep Learning models

Related tags

Overview

JAX Models

About The Project

Prerequisites

Installation

Usage

Contributing

License

Contact

You might also like...

Very deep VAEs in JAX/Flax

Conservative Q Learning for Offline Reinforcement Reinforcement Learning in JAX

PyTorch implementations of neural network models for keyword spotting

Unofficial implementation of Proxy Anchor Loss for Deep Metric Learning

Time-series-deep-learning - Developing Deep learning LSTM, BiLSTM models, and NeuralProphet for multi-step time-series forecasting of stock price.

FedJAX is a library for developing custom Federated Learning (FL) algorithms in JAX.

Objax Apache-2Objax (🥉19 · ⭐ 580) - Objax is a machine learning framework that provides an Object.. Apache-2 jax

Plug-n-Play Reinforcement Learning in Python with OpenAI Gym and JAX

JAX code for the paper "Control-Oriented Model-Based Reinforcement Learning with Implicit Differentiation"

Comments

Missing Axis Swap in ExtractPatches and MergePatches

fix convnext to make it work with jax.jit

Releases(v0.5-van)

v0.5-van(Feb 27, 2022)

v0.4-cait(Feb 18, 2022)

v0.3-pvit(Feb 11, 2022)

v0.2-convnext(Feb 6, 2022)

v0.1-swin(Jan 24, 2022)

Owner

The first dataset of composite images with rationality score indicating whether the object placement in a composite image is reasonable.

Neural Network to colorize grayscale images

Improving Contrastive Learning by Visualizing Feature Transformation, ICCV 2021 Oral

Repository containing detailed experiments related to the paper "Memotion Analysis through the Lens of Joint Embedding".

Code for SALT: Stackelberg Adversarial Regularization, EMNLP 2021.

Invariant Causal Prediction for Block MDPs

FastyAPI is a Stack boilerplate optimised for heavy loads.

Classification models 1D Zoo - Keras and TF.Keras

Scalable, Portable and Distributed Gradient Boosting (GBDT, GBRT or GBM) Library, for Python, R, Java, Scala, C++ and more. Runs on single machine, Hadoop, Spark, Dask, Flink and DataFlow

Really awesome semantic segmentation

Puzzle-CAM: Improved localization via matching partial and full features.

A repo with study material, exercises, examples, etc for Devnet SPAUTO

The missing CMake project initializer

This is the official implement of paper "ActionCLIP: A New Paradigm for Action Recognition"

EfficientNetV2-with-TPU - Cifar-10 case study

An official source code for "Augmentation-Free Self-Supervised Learning on Graphs"

Semi-Supervised 3D Hand-Object Poses Estimation with Interactions in Time

Simple Text-Generator with OpenAI gpt-2 Pytorch Implementation

Multimodal Co-Attention Transformer (MCAT) for Survival Prediction in Gigapixel Whole Slide Images

Adversarial Attacks on Probabilistic Autoregressive Forecasting Models.