High-Resolution 3D Assets Generation with Large Scale Diffusion Models
Automatically find issues in image datasets
Fast image augmentation library and an easy-to-use wrapper
This repo contains the code for 1D tokenizer and generator
ComfyUI wrapper nodes for HunyuanVideo
File and Image Management Application for django
Native and Compact Structured Latents for 3D Generation
AV1 Image File Format Specification - ISO-BMFF/HEIF derivative
Implementation of Imagen, Google's Text-to-Image Neural Network
Reverse engineering Gemini's SynthID detection
Seamlessly extend your preferred base images to be Lambda compatible
Reference PyTorch implementation and models for DINOv3
A Customizable Image-to-Video Model based on HunyuanVideo
An open-source photo thumbnail service by globo.com
Chinese and English multimodal conversational language model
AI PPT Track Terminator, the strongest PPT Skill ever
Official MiniMax Model Context Protocol (MCP) server
A neural network that transforms a design mock-up into static websites
State-of-the-art Image & Video CLIP, Multimodal Large Language Models
Train machine learning models within Docker containers
Tencent Hunyuan Multimodal diffusion transformer (MM-DiT) model
An open source implementation of CLIP
Blender addons to make the bridge between Blender and geographic data
[CVPR 2026 Oral] VGGT Omega
Essential nodes that are weirdly missing from ComfyUI core