SUPIR upscaling wrapper for ComfyUI
AI Image Upscaler & Enhancer
High-Resolution 3D Assets Generation with Large Scale Diffusion Models
Temporal-Consistent Diffusion Model for Real-World Video
AI tool that removes hardcoded subtitles and text from videos locally
IPTV live stream source automatic update tool
Official Python inference and LoRA trainer package
Official inference repo for FLUX.2 models
Cross platform GUI tool for downloading videos from Bilibili sites
Package manager based on libdnf and libsolv. Replaces YUM
Synthesizing and manipulating 2048x1024 images with conditional GANs
Repo for SeedVR2 & SeedVR
Recovering the Visual Space from Any Views
This repository contains the official implementation of FastVLM
OCRmyPDF adds an OCR text layer to scanned PDF files
A lossless video/GIF/image upscaler achieved with waifu2x, Anime4K
Foundational video generation model with 13.6B parameters
High-Resolution Image Synthesis with Latent Diffusion Models
Pure Python FFmpeg-based live video / audio streaming to YouTube
Knowledge Graph Generation from Any Text
MiniMax H3 is a general-purpose, omni-modal generative system
Reverse engineering Gemini's SynthID detection
GPT4V-level open-source multi-modal model based on Llama3-8B
Stable Diffusion web UI
Sharp Monocular View Synthesis in Less Than a Second