MII makes low-latency and high-throughput inference possible
The standard data-centric AI package for data quality and ML
Build cross-modal and multimodal applications on the cloud
Chinese and English multimodal conversational language model
A Codex Skill for generating high-density, editable PowerPoints
LISA: Reasoning Segmentation via Large Language Model
Skywork-R1V is an advanced multimodal AI model series
Refer and Ground Anything Anywhere at Any Granularity
Stable Diffusion WebUI optimized for AMD GPUs with editing tools
Language modeling in a sentence representation space
High-Resolution Image Synthesis with Latent Diffusion Models
Plug-n-play module turning text-to-image models into animation
Implementation of Phenaki Video, which uses Mask GIT
Run GGUF models easily with a UI or API. One File. Zero Install.
Mice speech to text with MX Cinnamon OS ISO
A Python application to add watermarks (text or image) to PDF files
Open source demo platform where you can easily showcase your AI models
Autoregressive Model Beats Diffusion
AI-powered tool to quickly remove watermarks from images flawlessly
Overcoming Data Limitations for High-Quality Video Diffusion Models
dashAI: an interactive platform for training, evaluating and deploying
A Pioneering Open-Source Alternative to GPT-4o
ktrain is a Python library that makes deep learning AI more accessible
Chat & pretrained large vision language model
Towards Real-World Vision-Language Understanding