Z80-μLM is a 2-bit quantized language model
Python tool for browser-based interactive data apps in one file
Cache-Augmented Generation: A Simple, Efficient Alternative to RAG
Convert TensorFlow, Keras, Tensorflow.js and Tflite models to ONNX
Library for efficiently connecting and optimizing teams of AI agents
AutoAgent: Fully-Automated and Zero-Code LLM Agent Framework
An MLOps framework to package, deploy, monitor and manage models
Train machine learning models within Docker containers
A fast TTS architecture with conditional flow matching
A solution to build and deploy MCP agents and applications
Building a Secure and Interoperable Future for AI-Driven Payments
Chat with your documents using local AI
OpenMMLab Model Deployment Framework
Embed images and sentences into fixed-length vectors
RNN with great LLM performance
A computer vision framework to create and deploy apps in minutes
Serve machine learning models within a Docker container
Tools for using Langchain with Prefect
CPU/GPU inference server for Hugging Face transformer models
A CLI tool/python module for generating images from text
Natural Language Processing Tutorial for Deep Learning Researchers
Source-to-source debuggable derivatives in pure Python