Video Object and Interaction Deletion
Buzz transcribes and translates audio offline
From Images to High-Fidelity 3D Assets
Awesome multilingual OCR toolkits based on PaddlePaddle
Open-source, high-performance AI model with advanced reasoning
Wan2.1: Open and Advanced Large-Scale Video Generative Model
The most powerful local music generation model
Wan2.2: Open and Advanced Large-Scale Video Generative Model
AlphaFold 3 inference pipeline
Official inference repo for FLUX.1 models
Official Python inference and LoRA trainer package
Generating Immersive, Explorable, and Interactive 3D Worlds
Native and Compact Structured Latents for 3D Generation
An experimental version of DeepSeek model
Fast stable diffusion on CPU and AI PC
A Family of Open Sourced Music Foundation Models
Agentic, Reasoning, and Coding (ARC) foundation models
Advanced language and coding AI model
1B text generation model based on the HRM architecture
Industrial-level controllable zero-shot text-to-speech system
LTX-Video Support for ComfyUI
code for Mesh R-CNN, ICCV 2019
Global weather forecasting model using graph neural networks and JAX
Open-source multi-speaker long-form text-to-speech model
State-of-the-art TTS model under 25MB