iOS/Android image picker with support for camera, video, etc.
Open Source Differentiable Computer Vision Library
An image processing library written entirely in JavaScript for Node
Multimodal-Driven Architecture for Customized Video Generation
Tencent Hunyuan Multimodal diffusion transformer (MM-DiT) model
State-of-the-art Image & Video CLIP, Multimodal Large Language Models
Ready-to-use OCR with 80+ supported languages
AutoGluon: AutoML for Image, Text, and Tabular Data
State-of-the-art Machine Learning for Pytorch, TensorFlow, and JAX
Spring AI Alibaba examples for building and testing AI apps
A distributed system for embedding-based vector retrieval
Jittor is a high-performance deep learning framework
Chinese and English multimodal conversational language model
Plug-n-play module turning text-to-image models into animation
Usable Implementation of "Bootstrap Your Own Latent" self-supervised
dashAI: an interactive platform for training, evaluating and deploying
Implements weak-to-strong learning for training stronger ML models
OpenMMLab Model Deployment Framework
A Python library for turning text quotes into graphical images
Text-conditional image generation model based on OpenAI's unCLIP
Node.js module for rendering pdf pages to images, svgs and HTML files
A CLI tool/python module for generating images from text
Next-generation platform for object detection and segmentation
A python module for hyperspectral image processing