Pythonic tool for running machine-learning/high performance workflows
Python module that helps you build complex pipelines of batch jobs
Streamline your ML workflow
Orange: Interactive data analysis
Scale your Pandas workflows by changing a single line of code
matplotlib: plotting with Python
Data integration platform for ELT pipelines from APIs, databases
Concurrent Python made simple
Monitor the stability of a Pandas or Spark dataframe
Training data (data labeling, annotation, workflow) for all data types
CKAN is an open-source DMS for powering data hubs
Build beautiful web-based analytic apps, no JavaScript required
Easy integration with Athena, Glue, Redshift, Timestream, Neptune
An orchestration platform for the development, production
Synthetic data generators for structured and unstructured text
Light-weight, flexible, expressive statistical data testing library
Make your own running home page
Benchmarking synthetic data generation methods
Train machine learning models within Docker containers
Positron, a next-generation data science IDE
Always know what to expect from your data
Docker image used to run data processing workloads
Data science on data without acquiring a copy
Data Analysis, Simulations and Visualization on the Sphere
Installable / Portable Python Distribution for Everyone.