Visualize and compare datasets, target values and associations
Python data, Leaflet.js maps
Open-source data observability for analytics engineers
Panda-Helper: Data profiling utility for Pandas DataFrames and Series
Light-weight, flexible, expressive statistical data testing library
The open standard for data logging
Main repository for Vispy
Recap tracks and transform schemas across your whole application
Python Stream Processing
Survival analysis in Python
Uncover insights, surface problems, monitor, and fine tune your LLM
Best practices on recommendation systems
tensorboard for pytorch (and chainer, mxnet, numpy, etc.)
An open source multi-tool for exploring and publishing data
Python module that helps you build complex pipelines of batch jobs
Positron, a next-generation data science IDE
Make your own running home page
High-Performance Symbolic Regression in Python and Julia
A curated list of data mining papers about fraud detection
Kubeflow’s superfood for Data Scientists
Pythonic tool for running machine-learning/high performance workflows
The standard data-centric AI package for data quality and ML
Benchmarking synthetic data generation methods
Metadata and data identification tool and Python library
Create HTML profiling reports from pandas DataFrame objects