Buzz transcribes and translates audio offline
MiniMax H3 is a general-purpose, omni-modal generative system
Wan2.2: Open and Advanced Large-Scale Video Generative Model
Awesome multilingual OCR toolkits based on PaddlePaddle
The most powerful local music generation model
Open-source, high-performance AI model with advanced reasoning
Run Bonsai (1-bit) and Ternary-Bonsai language models locally
Powerful AI language model (MoE) optimized for efficiency/performance
Official inference repo for FLUX.1 models
Wan2.1: Open and Advanced Large-Scale Video Generative Model
Official Python inference and LoRA trainer package
From Images to High-Fidelity 3D Assets
An experimental version of DeepSeek model
Code for running inference and finetuning with SAM 3 model
Native and Compact Structured Latents for 3D Generation
Python bindings for llama.cpp
GLM-4.5: Open-source LLM for intelligent agents by Z.ai
High-Resolution 3D Assets Generation with Large Scale Diffusion Models
Agentic, Reasoning, and Coding (ARC) foundation models
Fast stable diffusion on CPU and AI PC
Advanced language and coding AI model
Official inference repo for FLUX.2 models
Python inference and LoRA trainer package for the LTX-2 audio–video
Lets make video diffusion practical
Reference PyTorch implementation and models for DINOv3