Essential nodes that are weirdly missing from ComfyUI core
Official inference repo for FLUX.2 models
CLIP, Predict the most relevant text snippet given an image
Native and Compact Structured Latents for 3D Generation
Java interface to OpenCV, FFmpeg, and more
Capable of understanding text, audio, vision, video
ImageBind One Embedding Space to Bind Them All
Open-source evaluation toolkit of large multi-modality models (LMMs)
Automate your apps, games, and Android emulators
Implementation of "MobileCLIP" CVPR 2024
Implementation of Denoising Diffusion Probabilistic Model in Pytorch
Tensor search for humans
A skill that turn any brand into a scrollable 3D world
Fast and lightweight DNS proxy as ad-blocker for local network
The data structure for multimodal data
Multimodal embedding and reranking models built on Qwen3-VL
Qwen3-omni is a natively end-to-end, omni-modal LLM
Aligns and stacks astronomical images using point-set registration.
Native UI testing / controlling with node
Increase image resolution by eliminating atmospheric distortion
fast, minimal, portable image browser
Frontend emulator launcher displaying box art image of associated ROMs
Matches OpenScience Observatories images with astronomical catalogs
Free batch downloader for image, wallpaper, video, audio, document,
Free Animal Tracking Software