Aider is AI pair programming in your terminal
ComfyUI wrapper nodes for WanVideo and related models
InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System
Unified Multimodal Understanding and Generation Models
Fast-stable-diffusion + DreamBooth
Qwen2.5-VL is the multimodal large language model series
"Big Model" trains a visual multimodal VLM with 26M parameters
Implementation of "MobileCLIP" CVPR 2024
Fast multimodal LLM for real-time voice interaction and AI apps
Open source personal AI Assistant for Linux, Windows and Mac
Let your agent control your phone
An open sourced end-to-end VLM-based GUI Agent
Multimodal embedding and reranking models built on Qwen3-VL
Foundational video generation model with 13.6B parameters
21 Lessons, Get Started Building with Generative AI
High-Resolution Image Synthesis with Latent Diffusion Models
Stable Diffusion built-in to Blender
Generate Any 3D Scene in Seconds
95% token savings. 155x faster queries. 16 languages
Official Python inference and LoRA trainer package
Pretrained model hub for Keras 3
Sample code and notebooks for Generative AI on Google Cloud
Open-Sora: Democratizing Efficient Video Production for All
Phi-3.5 for Mac: Locally-run Vision and Language Models
Real-World Centric Foundation GUI Agents