liquid_scikit_learn

Scikit learn library models to account for data and concept drift.

This python library focuses on solving data drift and concept drift in the industry to minimize retraining of the models regularly. After inspired about the capabilities of neurons in octopus tentacles, which they interact and adapt directly with the environment without their central nervous system. I designed the weights for these models in the similar way where they train on input and experience. Instead of calculating weights based on minimizing the loss function, derivatives of weights are calculated. ( Hasani Chen). This library also provides model expiration details at a feature level. This could help in finding the features that model has hard time adjusting.

This library adapts concepts from Nueral ODE for scikit-learn. The models in this librabry calculate the derivatives of weights instead of weights as in standard scikit-learn librabry.

There are two training phases, the first one is a standard scikit learn model that provides predictions and weights for each feature. Typically, in standard ML models, training data is sent in batches and inferences can be done real time and in batch. In this scenario for the second training phase, input data is sent in semi batches and model adapts with changing data drift and concept drift with time. The second training phase along with changing weights it provides decay rate for each weight, contribution from data drift and concept drift and model failure parameters.

For example, suppose we train three months of data in the first training phase for the model to understand patterns with its provided inputs and outputs. In the second phase of training, we send weekly batches of inputs and outputs to make the model to adapt to changes in data and output that typically changes with customer behavior. I will make efforts to extend this library for unsupervised learning also. Currently liquid logistic regression is available with limited parameter optimization.

To use this librabry for now, git clone the librarby and give path to the librarby.

To use standard logistic regression

from liquid_scikit_learn.liquid_logistic_regression import logistic_regression

To use liquid logistic regression

from liquid_scikit_learn.liquid_logistic_regression import liquid_logistic_regression

To get model expiration details at a feature level

from liquid_scikit_learn.liquid_logistic_regression import model_failure

Scikit learn library models to account for data and concept drift.

Related tags

Overview

liquid_scikit_learn

Owner

Skoot is a lightweight python library of machine learning transformer classes that interact with scikit-learn and pandas.

A fast, distributed, high performance gradient boosting (GBT, GBDT, GBRT, GBM or MART) framework based on decision tree algorithms, used for ranking, classification and many other machine learning tasks.

Arquivos do curso online sobre a estatística voltada para ciência de dados e aprendizado de máquina.

A simple application that calculates the probability distribution of a normal distribution

pymc-learn: Practical Probabilistic Machine Learning in Python

Mesh TensorFlow: Model Parallelism Made Easier

Machine Learning Study 혼자 해보기

A Python Automated Machine Learning tool that optimizes machine learning pipelines using genetic programming.

FLAML is a lightweight Python library that finds accurate machine learning models automatically, efficiently and economically

Flightfare-Prediction - It is a Flightfare Prediction Web Application Using Machine learning,Python and flask

Distributed Tensorflow, Keras and PyTorch on Apache Spark/Flink & Ray

Credit Card Fraud Detection, used the credit card fraud dataset from Kaggle

Using Logistic Regression and classifiers of the dataset to produce an accurate recall, f-1 and precision score

whylogs: A Data and Machine Learning Logging Standard

A data preprocessing package for time series data. Design for machine learning and deep learning.

The Fuzzy Labs guide to the universe of open source MLOps

A comprehensive set of fairness metrics for datasets and machine learning models, explanations for these metrics, and algorithms to mitigate bias in datasets and models.

A statistical library designed to fill the void in Python's time series analysis capabilities, including the equivalent of R's auto.arima function.

MIT-Machine Learning with Python–From Linear Models to Deep Learning

Predico Disease Prediction system based on symptoms provided by patient- using Python-Django & Machine Learning