Fast stable diffusion on CPU and AI PC
Buzz transcribes and translates audio offline
Open-source, high-performance AI model with advanced reasoning
Convert Google Gemini web into OpenAI-compatible API
Wan2.1: Open and Advanced Large-Scale Video Generative Model
Tongyi Deep Research, the Leading Open-source Deep Research Agent
Advanced language and coding AI model
From Images to High-Fidelity 3D Assets
The most powerful local music generation model
AlphaFold 3 inference pipeline
Wan2.2: Open and Advanced Large-Scale Video Generative Model
Official inference repo for FLUX.1 models
Awesome multilingual OCR toolkits based on PaddlePaddle
Generating Immersive, Explorable, and Interactive 3D Worlds
Native and Compact Structured Latents for 3D Generation
AI PPT Track Terminator, the strongest PPT Skill ever
Official Python inference and LoRA trainer package
Text and image to video generation: CogVideoX and CogVideo
High-Resolution Image Synthesis with Latent Diffusion Models
Industrial-level controllable zero-shot text-to-speech system
An experimental version of DeepSeek model
A SOTA open-source image editing model
State-of-the-art TTS model under 25MB
Agentic, Reasoning, and Coding (ARC) foundation models
gpt-oss-120b and gpt-oss-20b are two open-weight language models