Turn WiFi signals into real-time human pose estimation and detection
Instant voice cloning by MIT and MyShell. Audio foundation model
Export Django monitoring metrics for Prometheus.io
MOSS-TTS-Nano is an open-source multilingual tiny speech generation
Foundation model for image generation
Learning agent trained in a diffusion world model
InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System
Code and models for ICML 2024 paper, NExT-GPT
A lightweight text-to-speech model with zero-shot voice cloning
Chinese XLNet pre-trained model
StreamSpeech is a seamless model for offline speech recognition
DeepEP: an efficient expert-parallel communication library
Easy-to-use,Modular and Extendible package of deep-learning models
Making large AI models cheaper, faster and more accessible
An MCP server that provides fast file searching capabilities
CLI tool to build, test, debug, and deploy Serverless applications
Gemma open-weight LLM library, from Google DeepMind
Agentic, Reasoning, and Coding (ARC) foundation models
Tool for producing high quality forecasts for time series data
Multimodal AI chat app with dynamic conversation routing
Version-independent Codex instruction deployment
No-code in the front, Python in the back. An open-source framework
PandasAI is a Python library that integrates generative AI
BertViz: Visualize Attention in NLP Models (BERT, GPT2, BART, etc.)
The open-source tool for building high-quality datasets