Open Source Differentiable Computer Vision Library
Multimodal-Driven Architecture for Customized Video Generation
Tencent Hunyuan Multimodal diffusion transformer (MM-DiT) model
State-of-the-art Image & Video CLIP, Multimodal Large Language Models
Ready-to-use OCR with 80+ supported languages
AutoGluon: AutoML for Image, Text, and Tabular Data
State-of-the-art Machine Learning for Pytorch, TensorFlow, and JAX
Spring AI Alibaba examples for building and testing AI apps
Jittor is a high-performance deep learning framework
Chinese and English multimodal conversational language model
Plug-n-play module turning text-to-image models into animation
Usable Implementation of "Bootstrap Your Own Latent" self-supervised
dashAI: an interactive platform for training, evaluating and deploying
Implements weak-to-strong learning for training stronger ML models
OpenMMLab Model Deployment Framework
Generate Harmonious Colors Freely.
A cross-platform GUI automation Python module for human beings
A Python library for turning text quotes into graphical images
Text-conditional image generation model based on OpenAI's unCLIP
A CLI tool/python module for generating images from text
Next-generation platform for object detection and segmentation
A python module for hyperspectral image processing
Deal with bad samples in your dataset dynamically
A fully customisable subclass of the native UIControl