Spark-TTS Inference Code
Docker image used to run data processing workloads
Jupyter magics and kernels for working with remote Spark clusters
A unified interface for distributed computing
Python Stream Processing
Monitor the stability of a Pandas or Spark dataframe
Scalable machine learning for time series forecasting
Dataproc templates and pipelines for solving simple in-cloud data task
MLOps simplified. From ML Pipeline ⇨ Data Product without the hassle
Open source platform for the machine learning lifecycle
DoWhy is a Python library for causal inference
Fast forecasting with statistical and econometric models
Unified Model Serving Framework
NumPy aware dynamic Python compiler using LLVM
A Python framework for creating reproducible, maintainable code
A lightweight data processing framework built on DuckDB and 3FS
Distributed training framework for TensorFlow, Keras, PyTorch, etc.
Source code accompanying book: Data Science on the GCP
Distributed Deep learning with Keras & Spark
Reading OpenStreetMap Pbf files.
TensorFlowOnSpark brings TensorFlow programs to Apache Spark clusters
A Deep Learning Recommender System
A recommender system for discovering GitHub repos
Python Helper library for Jupyter Notebooks