Convert AI papers to GUI
Large-language-model & vision-language-model based on Linear Attention
Chinese and English multimodal conversational language model
High-Resolution Image Synthesis with Latent Diffusion Models
Open source framework for deep learning satellite and aerial imagery
Scientific Visualisation Made Easy
Run GGUF models easily with a UI or API. One File. Zero Install.
Implementation of Phenaki Video, which uses Mask GIT
Plug-n-play module turning text-to-image models into animation
Open source demo platform where you can easily showcase your AI models
Tooling for the Common Objects In 3D dataset
Graphical User Interface Face Anonymization Tool
Visual Automation IDE — automate anything you see on screen
A Customizable Image-to-Video Model based on HunyuanVideo
Autoregressive Model Beats Diffusion
Powerful open source image generation model
Usable Implementation of "Bootstrap Your Own Latent" self-supervised
A Python application to add watermarks (text or image) to PDF files
Mice speech to text with MX Cinnamon OS ISO
Overcoming Data Limitations for High-Quality Video Diffusion Models
Open-source AI video pipeline, fully automated with MCP
dashAI: an interactive platform for training, evaluating and deploying
Multi-user UI for managing and running Stable Diffusion workflows tool
It's possible for machines to become self-aware.
Easy Docker setup for Stable Diffusion with user-friendly UI