A Powerful Serverless Analysis Toolkit That Takes Trial And Error Out of Machine Learning Projects

Last update: Jan 02, 2023

Overview

KXY: A Seemless API to 10x The Productivity of Machine Learning Engineers

Documentation

https://www.kxy.ai/reference/

Installation

From PyPi:

pip install kxy

From GitHub:

git clone https://github.com/kxytechnologies/kxy-python.git & cd ./kxy-python & pip install .

Authentication

All heavy-duty computations are run on our serverless infrastructure and require an API key. To configure the package with your API key, run

kxy configure

and follow the instructions. To get an API key you need an account; you can sign up for a free trial here. You'll then be automatically given an API key which you can find here.

KXY is free for academic use.

Docker

The Docker image kxytechnologies/kxy has been built for your convenience, and comes with anaconda, auto-sklearn, and the kxy package.

To start a Jupyter Notebook server from a sandboxed Docker environment, run

&& /opt/conda/bin/jupyter notebook --notebook-dir=/opt/notebooks --ip='*' --port=8888 --no-browser --allow-root --NotebookApp.token=''" ">

docker run -i -t -p 5555:8888 kxytechnologies/kxy:latest /bin/bash -c "kxy configure 
   
     && /opt/conda/bin/jupyter notebook --notebook-dir=/opt/notebooks --ip='*' --port=8888 --no-browser --allow-root --NotebookApp.token=''
    "

where you should replace with your API key and navigate to http://localhost:5555 in your browser. This docker environment comes with all examples available on the documentation website.

To start a Jupyter Notebook server from an existing directory of notebooks, run

&& /opt/conda/bin/jupyter notebook --notebook-dir=/opt/notebooks --ip='*' --port=8888 --no-browser --allow-root --NotebookApp.token=''" ">

docker run -i -t --mount src=</path/to/your/local/dir>,target=/opt/notebooks,type=bind -p 5555:8888 kxytechnologies/kxy:latest /bin/bash -c "kxy configure 
   
     && /opt/conda/bin/jupyter notebook --notebook-dir=/opt/notebooks --ip='*' --port=8888 --no-browser --allow-root --NotebookApp.token=''
    "

where you should replace with the path to your local notebook folder and navigate to http://localhost:5555 in your browser.

Other Programming Language

We plan to release friendly API client in more programming language.

In the meantime, you can directly issue requests to our RESTFul API using your favorite programming language.

Kubeflow is a machine learning (ML) toolkit that is dedicated to making deployments of ML workflows on Kubernetes simple, portable, and scalable.

SDK: Overview of the Kubeflow pipelines service Kubeflow is a machine learning (ML) toolkit that is dedicated to making deployments of ML workflows on

3.1k Jan 6, 2023

Model Validation Toolkit is a collection of tools to assist with validating machine learning models prior to deploying them to production and monitoring them after deployment to production.

25 Dec 28, 2022

A machine learning toolkit dedicated to time-series data

tslearn The machine learning toolkit for time series analysis in Python Section Description Installation Installing the dependencies and tslearn Getti

2.3k Jan 5, 2023

A machine learning toolkit dedicated to time-series data

tslearn The machine learning toolkit for time series analysis in Python Section Description Installation Installing the dependencies and tslearn Getti

2.3k Dec 29, 2022

Kats is a toolkit to analyze time series data, a lightweight, easy-to-use, and generalizable framework to perform time series analysis.

Kats, a kit to analyze time series data, a lightweight, easy-to-use, generalizable, and extendable framework to perform time series analysis, from understanding the key statistics and characteristics, detecting change points and anomalies, to forecasting future trends.

4.1k Dec 29, 2022

A mindmap summarising Machine Learning concepts, from Data Analysis to Deep Learning.

5.7k Dec 30, 2022

A library of extension and helper modules for Python's data analysis and machine learning libraries.

Mlxtend (machine learning extensions) is a Python library of useful tools for the day-to-day data science tasks. Sebastian Raschka 2014-2021 Links Doc

4.2k Dec 29, 2022

A Python Automated Machine Learning tool that optimizes machine learning pipelines using genetic programming.

Master status: Development status: Package information: TPOT stands for Tree-based Pipeline Optimization Tool. Consider TPOT your Data Science Assista

8.9k Jan 9, 2023

Python Extreme Learning Machine (ELM) is a machine learning technique used for classification/regression tasks.

Python Extreme Learning Machine (ELM) Python Extreme Learning Machine (ELM) is a machine learning technique used for classification/regression tasks.

84 Nov 25, 2022

Comments

error in import kxy

Hi, After installing the kxy package and configuring the API key, the import kxy shows the error below:

.../python3.9/site-packages/kxy/pfs/pfs_selector.py in <module>
      6 import numpy as np
      7 
----> 8 import tensorflow as tf
      9 from tensorflow.keras.callbacks import EarlyStopping, TerminateOnNaN
     10 from tensorflow.keras.optimizers import Adam

ModuleNotFoundError: No module named 'tensorflow'

what version of tensorflow is needed for kxy to work?

opened by zeydabadi 2

generate_features Documentation?

Is there any documentation on how to use the generate_features function? It doesn't appear in the documentation and I can't find it in the github. e.g. how to use the entity column, how to format time-series data in advance for it, etc'. Thanks!

opened by ddofer 1
error kxy.data_valuation

Hi, After running chievable_performance_df = X_train_reduced.kxy.data_valuation(target_column='state', problem_type='classification', include_mutual_information=True, anonymize=True) I get the following error and the function does not return anything: `During handling of the above exception, another exception occurred:

Traceback (most recent call last): File "/usr/lib/python3.9/asyncio/tasks.py", line 258, in __step result = coro.throw(exc) File "/home/lucy/Downloads/general/lib/python3.9/site-packages/tornado/websocket.py", line 1104, in wrapper raise WebSocketClosedError() tornado.websocket.WebSocketClosedError Task exception was never retrieved future: <Task finished name='Task-46004' coro=<WebSocketProtocol13.write_message..wrapper() done, defined at /home/lucy/Downloads/general/lib/python3.9/site-packages/tornado/websocket.py:1100> exception=WebSocketClosedError()> Traceback (most recent call last): File "/home/lucy/Downloads/general/lib/python3.9/site-packages/tornado/websocket.py", line 1102, in wrapper await fut File "/usr/lib/python3.9/asyncio/tasks.py", line 328, in __wakeup future.result() tornado.iostream.StreamClosedError: Stream is closed `

opened by zeydabadi 0

Releases(v1.4.10)

v1.4.10(Apr 25, 2022)
Change Log

v.1.4.10 Changes

Added a function to construct features derived from PFS mutual information estimation that should be expected to be linearly related to the target.

Fixed a global name conflict in kxy.learning.base_learners.

v.1.4.9 Changes

Change the activation function used by PFS from ReLU to switch/SILU.

Leaving it to the user to set the logging level.

v.1.4.8 Changes

Froze the versions of all python packages in the docker file.

v.1.4.7 Changes

Changes related to optimizing Principal Feature Selection.

Made it easy to change PFS' default learning parameters.

Changed PFS' default learning parameters (learning rate is now 0.005 and epsilon 1e-04)

Adding a seed parameter to PFS' fit for reproducibility.

To globally change the learning rate to 0.003, change Adam's epsilon to 1e-5, and the number of epochs to 25, do

from kxy.misc.tf import set_default_parameter set_default_parameter('lr', 0.003) set_default_parameter('epsilon', 1e-5) set_default_parameter('epochs', 25)

To change the number epochs for a single iteration of PFS, use the epochs argument of the fit method of your PFS object. The fit method now also has a seed parameter you may use to make the PFS implementation deterministic.

Example:

from kxy.pfs import PFS selector = PFS() selector.fit(x, y, epochs=25, seed=123)

Alternatively, you may also use the kxy.misc.tf.set_seed method to make PFS deterministic.

v.1.4.6 Changes

Minor PFS improvements.

Adding more (robust) mutual information loss functions.

Exposing the learned total mutual information between principal features and target as an attribute of PFS.

Exposing the number of epochs as a parameter of PFS' fit.

Source code(tar.gz)
Source code(zip)
v1.4.9(Apr 12, 2022)
Change Log

v.1.4.9 Changes

Change the activation function used by PFS from ReLU to switch/SILU.

Leaving it to the user to set the logging level.

v.1.4.8 Changes

Froze the versions of all python packages in the docker file.

v.1.4.7 Changes

Changes related to optimizing Principal Feature Selection.

Made it easy to change PFS' default learning parameters.

Changed PFS' default learning parameters (learning rate is now 0.005 and epsilon 1e-04)

Adding a seed parameter to PFS' fit for reproducibility.

To globally change the learning rate to 0.003, change Adam's epsilon to 1e-5, and the number of epochs to 25, do

from kxy.misc.tf import set_default_parameter set_default_parameter('lr', 0.003) set_default_parameter('epsilon', 1e-5) set_default_parameter('epochs', 25)

To change the number epochs for a single iteration of PFS, use the epochs argument of the fit method of your PFS object. The fit method now also has a seed parameter you may use to make the PFS implementation deterministic.

Example:

from kxy.pfs import PFS selector = PFS() selector.fit(x, y, epochs=25, seed=123)

Alternatively, you may also use the kxy.misc.tf.set_seed method to make PFS deterministic.

v.1.4.6 Changes

Minor PFS improvements.

Adding more (robust) mutual information loss functions.

Exposing the learned total mutual information between principal features and target as an attribute of PFS.

Exposing the number of epochs as a parameter of PFS' fit.

Source code(tar.gz)
Source code(zip)
v1.4.8(Apr 11, 2022)
Change Log

v.1.4.8 Changes

Froze the versions of all python packages in the docker file.

v.1.4.7 Changes

Changes related to optimizing Principal Feature Selection.

Made it easy to change PFS' default learning parameters.

Changed PFS' default learning parameters (learning rate is now 0.005 and epsilon 1e-04)

Adding a seed parameter to PFS' fit for reproducibility.

To globally change the learning rate to 0.003, change Adam's epsilon to 1e-5, and the number of epochs to 25, do

from kxy.misc.tf import set_default_parameter set_default_parameter('lr', 0.003) set_default_parameter('epsilon', 1e-5) set_default_parameter('epochs', 25)

To change the number epochs for a single iteration of PFS, use the epochs argument of the fit method of your PFS object. The fit method now also has a seed parameter you may use to make the PFS implementation deterministic.

Example:

from kxy.pfs import PFS selector = PFS() selector.fit(x, y, epochs=25, seed=123)

Alternatively, you may also use the kxy.misc.tf.set_seed method to make PFS deterministic.

v.1.4.6 Changes

Minor PFS improvements.

Adding more (robust) mutual information loss functions.

Exposing the learned total mutual information between principal features and target as an attribute of PFS.

Exposing the number of epochs as a parameter of PFS' fit.

Source code(tar.gz)
Source code(zip)
v1.4.7(Apr 10, 2022)
Change Log

v.1.4.7 Changes

Changes related to optimizing Principal Feature Selection.

Made it easy to change PFS' default learning parameters.

Changed PFS' default learning parameters (learning rate is now 0.005 and epsilon 1e-04)

Adding a seed parameter to PFS' fit for reproducibility.

To globally change the learning rate to 0.003, change Adam's epsilon to 1e-5, and the number of epochs to 25, do

from kxy.misc.tf import set_default_parameter set_default_parameter('lr', 0.003) set_default_parameter('epsilon', 1e-5) set_default_parameter('epochs', 25)

To change the number epochs for a single iteration of PFS, use the epochs argument of the fit method of your PFS object. The fit method now also has a seed parameter you may use to make the PFS implementation deterministic.

Example:

from kxy.pfs import PFS selector = PFS() selector.fit(x, y, epochs=25, seed=123)

Alternatively, you may also use the kxy.misc.tf.set_seed method to make PFS deterministic.

v.1.4.6 Changes

Minor PFS improvements.

Adding more (robust) mutual information loss functions.

Exposing the learned total mutual information between principal features and target as an attribute of PFS.

Exposing the number of epochs as a parameter of PFS' fit.

Source code(tar.gz)
Source code(zip)
v1.4.6(Apr 10, 2022)
Changes

Adding more (robust) mutual information loss functions.

Exposing the learned total mutual information between principal features and target as an attribute of PFS.

Exposing the number of epochs as a parameter of PFS' fit.

Source code(tar.gz)
Source code(zip)
v1.4.5(Apr 9, 2022)

Fixing some package incompatibilities.
Source code(tar.gz)
Source code(zip)
v1.4.4(Apr 8, 2022)

Adding Principal Feature Selection.
Source code(tar.gz)
Source code(zip)
v1.0.4(Jul 1, 2021)

Source code(tar.gz)
Source code(zip)
1.0.3(Mar 23, 2021)

Source code(tar.gz)
Source code(zip)
1.0.2(Mar 16, 2021)

Source code(tar.gz)
Source code(zip)
1.0.1(Mar 16, 2021)

Source code(tar.gz)
Source code(zip)
0.3.8(Jan 25, 2021)

Source code(tar.gz)
Source code(zip)
v0.3.5(Jan 21, 2021)

Source code(tar.gz)
Source code(zip)
v0.3.4(Dec 16, 2020)

Source code(tar.gz)
Source code(zip)
v0.3.2(Aug 14, 2020)

Adding regression root mean square error (RMSE) in the list of metrics whose achievable values we calculate.
Source code(tar.gz)
Source code(zip)
v0.3.1(Aug 7, 2020)

Source code(tar.gz)
Source code(zip)
v0.3.0(Aug 3, 2020)

Adding a maximum-entropy based classifier (kxy.MaxEntClassifier) and regressor (kxy.MaxEntRegressor) following the scikit-learn signature for fitting and predicting.

These models estimate the posterior mean E[u_y|x] and the posterior standard deviation sqrt(Var[u_y|x]) for any specific value of x, where the copula-uniform representations (u_y, u_x) follow the maximum-entropy distribution.

Predictions in the primal are derived from E[u_y|x].
Source code(tar.gz)
Source code(zip)
v0.2.0(Jun 25, 2020)
Regression analyses now fully support categorical variables.

Foundations for multi-output regressions are laid.

Categorical variables are now systematically encoded and treated as continuous, consistent with what's done at the learning stage.

Regression and classification are further normalized, and most the compute for classification problems now takes place on the API side, and should be considerably faster.

Source code(tar.gz)
Source code(zip)
v0.1.3(Jun 12, 2020)

Source code(tar.gz)
Source code(zip)
v0.1.2(Jun 12, 2020)

Source code(tar.gz)
Source code(zip)
v0.1.1(Jun 11, 2020)

Source code(tar.gz)
Source code(zip)
v0.0.18(May 26, 2020)

Source code(tar.gz)
Source code(zip)
v0.0.16(May 18, 2020)

Source code(tar.gz)
Source code(zip)
v0.0.15(May 18, 2020)

Source code(tar.gz)
Source code(zip)
v0.0.14(May 18, 2020)

Source code(tar.gz)
Source code(zip)
v0.0.13(May 16, 2020)

Source code(tar.gz)
Source code(zip)
v0.0.11(May 13, 2020)

Source code(tar.gz)
Source code(zip)
v0.0.10(May 11, 2020)

Source code(tar.gz)
Source code(zip)
v0.0.3(Apr 17, 2020)

Source code(tar.gz)
Source code(zip)
v0.0.2(Apr 17, 2020)

Source code(tar.gz)
Source code(zip)

Owner

KXY Technologies, Inc.

GitHub Repository https://kxy.ai

A simple guide to MLOps through ZenML and its various integrations.

ZenBytes Join our Slack Community and become part of the ZenML family Give the main ZenML repo a GitHub star to show your love ZenBytes is a series of

127 Dec 27, 2022

moDel Agnostic Language for Exploration and eXplanation

moDel Agnostic Language for Exploration and eXplanation Overview Unverified black box model is the path to the failure. Opaqueness leads to distrust.

1.2k Jan 04, 2023

Responsible AI Workshop: a series of tutorials & walkthroughs to illustrate how put responsible AI into practice

Responsible AI Workshop Responsible innovation is top of mind. As such, the tech industry as well as a growing number of organizations of all kinds in

9 Sep 14, 2022

OptaPy is an AI constraint solver for Python to optimize planning and scheduling problems.

OptaPy is an AI constraint solver for Python to optimize the Vehicle Routing Problem, Employee Rostering, Maintenance Scheduling, Task Assignment, School Timetabling, Cloud Optimization, Conference S

208 Dec 27, 2022

a distributed deep learning platform

Apache SINGA Distributed deep learning system http://singa.apache.org Quick Start Installation Examples Issues JIRA tickets Code Analysis: Mailing Lis

2.7k Jan 05, 2023

Unofficial pytorch implementation of the paper "Context Reasoning Attention Network for Image Super-Resolution (ICCV 2021)"

CRAN Unofficial pytorch implementation of the paper "Context Reasoning Attention Network for Image Super-Resolution (ICCV 2021)" This code doesn't exa

4 Nov 11, 2021

Programming assignments and quizzes from all courses within the Machine Learning Engineering for Production (MLOps) specialization offered by deeplearning.ai

Machine Learning Engineering for Production (MLOps) Specialization on Coursera (offered by deeplearning.ai) Programming assignments from all courses i

173 Jan 05, 2023

Machine learning algorithms implementation

Machine learning algorithms implementation This repository consisits of implementation of various machine learning algorithms. The algorithms implemen

1 Jan 03, 2022

Forecasting prices using Facebook/Meta's Prophet model

CryptoForecasting using Machine and Deep learning (Part 1) CryptoForecasting using Machine Learning The main aspect of predicting the stock-related da

1 Nov 27, 2021

A python fast implementation of the famous SVD algorithm popularized by Simon Funk during Netflix Prize

⚡ funk-svd funk-svd is a Python 3 library implementing a fast version of the famous SVD algorithm popularized by Simon Funk during the Neflix Prize co

171 Dec 19, 2022

mlpack: a scalable C++ machine learning library --

4.2k Jan 01, 2023

BentoML is a flexible, high-performance framework for serving, managing, and deploying machine learning models.

Model Serving Made Easy BentoML is a flexible, high-performance framework for serving, managing, and deploying machine learning models. Supports multi

4.4k Jan 04, 2023

A benchmark of data-centric tasks from across the machine learning lifecycle.

61 Dec 28, 2022

DistML is a Ray extension library to support large-scale distributed ML training on heterogeneous multi-node multi-GPU clusters

27 Aug 19, 2022

A statistical library designed to fill the void in Python's time series analysis capabilities, including the equivalent of R's auto.arima function.

pmdarima Pmdarima (originally pyramid-arima, for the anagram of 'py' + 'arima') is a statistical library designed to fill the void in Python's time se

1.3k Dec 22, 2022

Stats, linear algebra and einops for xarray

xarray-einstats Stats, linear algebra and einops for xarray ⚠️ Caution: This project is still in a very early development stage Installation To instal

30 Dec 28, 2022

Tangram makes it easy for programmers to train, deploy, and monitor machine learning models.

Tangram Website | Discord Tangram makes it easy for programmers to train, deploy, and monitor machine learning models. Run tangram train to train a mo

1.4k Jan 05, 2023

Kubeflow is a machine learning (ML) toolkit that is dedicated to making deployments of ML workflows on Kubernetes simple, portable, and scalable.

SDK: Overview of the Kubeflow pipelines service Kubeflow is a machine learning (ML) toolkit that is dedicated to making deployments of ML workflows on

3.1k Jan 06, 2023

Machine Learning e Data Science com Python

Machine Learning e Data Science com Python Arquivos do curso de Data Science e Machine Learning com Python na Udemy, cliqe aqui para acessá-lo. O prin

1 Jan 27, 2022

Cryptocurrency price prediction and exceptions in python

Cryptocurrency price prediction and exceptions in python This is a coursework on foundations of computing module Through this coursework i worked on m

1 Nov 07, 2021

A Powerful Serverless Analysis Toolkit That Takes Trial And Error Out of Machine Learning Projects

Related tags

Overview

KXY: A Seemless API to 10x The Productivity of Machine Learning Engineers

Documentation

Installation

Authentication

Docker

Other Programming Language

You might also like...

Kubeflow is a machine learning (ML) toolkit that is dedicated to making deployments of ML workflows on Kubernetes simple, portable, and scalable.

Model Validation Toolkit is a collection of tools to assist with validating machine learning models prior to deploying them to production and monitoring them after deployment to production.

A machine learning toolkit dedicated to time-series data

A machine learning toolkit dedicated to time-series data

Kats is a toolkit to analyze time series data, a lightweight, easy-to-use, and generalizable framework to perform time series analysis.

A mindmap summarising Machine Learning concepts, from Data Analysis to Deep Learning.

A library of extension and helper modules for Python's data analysis and machine learning libraries.

A Python Automated Machine Learning tool that optimizes machine learning pipelines using genetic programming.

Python Extreme Learning Machine (ELM) is a machine learning technique used for classification/regression tasks.

Comments

error in import kxy

generate_features Documentation?

error kxy.data_valuation

Releases(v1.4.10)

v1.4.10(Apr 25, 2022)

Change Log

v.1.4.10 Changes

v.1.4.9 Changes

v.1.4.8 Changes

v.1.4.7 Changes

v.1.4.6 Changes

v1.4.9(Apr 12, 2022)

Change Log

v.1.4.9 Changes

v.1.4.8 Changes

v.1.4.7 Changes

v.1.4.6 Changes

v1.4.8(Apr 11, 2022)

Change Log

v.1.4.8 Changes

v.1.4.7 Changes

v.1.4.6 Changes

v1.4.7(Apr 10, 2022)

Change Log

v.1.4.7 Changes

v.1.4.6 Changes

v1.4.6(Apr 10, 2022)

Changes

v1.4.5(Apr 9, 2022)

v1.4.4(Apr 8, 2022)

v1.0.4(Jul 1, 2021)

1.0.3(Mar 23, 2021)

1.0.2(Mar 16, 2021)

1.0.1(Mar 16, 2021)

0.3.8(Jan 25, 2021)

v0.3.5(Jan 21, 2021)

v0.3.4(Dec 16, 2020)

v0.3.2(Aug 14, 2020)

v0.3.1(Aug 7, 2020)

v0.3.0(Aug 3, 2020)

v0.2.0(Jun 25, 2020)

v0.1.3(Jun 12, 2020)

v0.1.2(Jun 12, 2020)

v0.1.1(Jun 11, 2020)

v0.0.18(May 26, 2020)

v0.0.16(May 18, 2020)

v0.0.15(May 18, 2020)

v0.0.14(May 18, 2020)

v0.0.13(May 16, 2020)

v0.0.11(May 13, 2020)

v0.0.10(May 11, 2020)

v0.0.3(Apr 17, 2020)

v0.0.2(Apr 17, 2020)

Owner

KXY Technologies, Inc.

A simple guide to MLOps through ZenML and its various integrations.

moDel Agnostic Language for Exploration and eXplanation

Responsible AI Workshop: a series of tutorials & walkthroughs to illustrate how put responsible AI into practice

OptaPy is an AI constraint solver for Python to optimize planning and scheduling problems.

a distributed deep learning platform