Python bindings for llama.cpp
Provides convenient access to the Anthropic REST API from any Python 3
Tencent Hunyuan Multimodal diffusion transformer (MM-DiT) model
Python SDK for Claude Agent
Multimodal-Driven Architecture for Customized Video Generation
AI cognitive-enhancement Skills based on Anthropic's J-space
Buzz transcribes and translates audio offline
Official inference repo for FLUX.1 models
MiniMax H3 is a general-purpose, omni-modal generative system
From Images to High-Fidelity 3D Assets
The most powerful local music generation model
Awesome multilingual OCR toolkits based on PaddlePaddle
Official Python inference and LoRA trainer package
State-of-the-art Image & Video CLIP, Multimodal Large Language Models
State-of-the-art TTS model under 25MB
Python inference and LoRA trainer package for the LTX-2 audio–video
Code for running inference and finetuning with SAM 3 model
Qwen's most powerful open-source image generation model
Open-source, high-performance AI model with advanced reasoning
Wan2.2: Open and Advanced Large-Scale Video Generative Model
Wan2.1: Open and Advanced Large-Scale Video Generative Model
Code for running inference with the SAM 3D Body Model 3DB
High-Fidelity and Controllable Generation of Textured 3D Assets
Native and Compact Structured Latents for 3D Generation
Models for object and human mesh reconstruction