Generating Immersive, Explorable, and Interactive 3D Worlds
Pokee Deep Research Model Open Source Repo
Understanding of AI Infra: Quantitative Analysis and System Design
Your Personal AI Assistant; easy to install, deploy on local or coud
StreamSpeech is a seamless model for offline speech recognition
Streamline your ML workflow
Multilingual Document Layout Parsing in a Single Vision-Language Model
Repo of Qwen2-Audio chat & pretrained large audio language model
Speech-AI-Forge is a project developed around TTS generation model
Open-source MCP server that gives your coding agent
A neural network that transforms a design mock-up into static websites
Build AI WhatsApp Bots with Pure Python
Open-source AI video pipeline, fully automated with MCP
One-click deployment (including offline integration package)
ChatGLM2-6B: An Open Bilingual Chat LLM
A state-of-the-art open visual language model
VITS2 backbone with multilingual-bert
Multi-Voice and Prompt-Controlled TTS Engine
Reverse engineered ChatGPT API
Python SDK/API for reverse engineered Google Bard
A webui for different audio related Neural Networks
RWKV for Chinese novel generation
Clone a voice in 5 seconds to generate arbitrary speech in real-time
DSTK - DataScience ToolKit for All of Us