Context data platform for building observable, self-learning AI agents
Dockerized FastAPI wrapper for Kokoro-82M text-to-speech model
code for Mesh R-CNN, ICCV 2019
A series of math-specific large language models of our Qwen2 series
Chatbot daemon that connects to your favorite chat services
Generate high-definition story short videos with one click using AI
A TTS model capable of generating ultra-realistic dialogue
Plug-and-play library to enable agents to call MCP and UTCP tools
Tongyi Deep Research, the Leading Open-source Deep Research Agent
A solution to build and deploy MCP agents and applications
The data structure for multimodal data
Toolkit for conversational AI
A multi-function Discord bot
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
High-Fidelity and Controllable Generation of Textured 3D Assets
State-of-the-art (SoTA) text-to-video pre-trained model
OCR expert VLM powered by Hunyuan's native multimodal architecture
RGBD video generation model conditioned on camera input
Automatically translates the text of a video based on a subtitle file
Qwen3-omni is a natively end-to-end, omni-modal LLM
MedicalGPT: Training Your Own Medical GPT Model with ChatGPT Training
A python library for easy manipulation and forecasting of time series
Capable of understanding text, audio, vision, video
The official implementation of RAPTOR
Run LLM prompts from your shell