Buzz transcribes and translates audio offline
Wan2.2: Open and Advanced Large-Scale Video Generative Model
Open-source, high-performance AI model with advanced reasoning
The most powerful local music generation model
Powerful AI language model (MoE) optimized for efficiency/performance
Official inference repo for FLUX.1 models
Wan2.1: Open and Advanced Large-Scale Video Generative Model
Awesome multilingual OCR toolkits based on PaddlePaddle
Run Bonsai (1-bit) and Ternary-Bonsai language models locally
Python bindings for llama.cpp
Native and Compact Structured Latents for 3D Generation
An experimental version of DeepSeek model
GLM-4.5: Open-source LLM for intelligent agents by Z.ai
Official inference repo for FLUX.2 models
A Family of Open Sourced Music Foundation Models
Python inference and LoRA trainer package for the LTX-2 audio–video
Accurate × Fast × Comprehensive
Official Python inference and LoRA trainer package
Qwen3-TTS is an open-source series of TTS models
Open-source multi-speaker long-form text-to-speech model
Fast stable diffusion on CPU and AI PC
Lets make video diffusion practical
Reference PyTorch implementation and models for DINOv3
Agentic, Reasoning, and Coding (ARC) foundation models
Text and image to video generation: CogVideoX and CogVideo