GeoAI: Artificial Intelligence for Geospatial Data
HunyuanVideo: A Systematic Framework For Large Video Generation Model
Advanced AI Explainability for computer vision
Official implementation of DreamCraft3D
Spring AI Alibaba examples for building and testing AI apps
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
Build cross-modal and multimodal applications on the cloud
An extensive node suite that enables ComfyUI to process 3D inputs
AI-data warehouse to enrich, transform and analyze unstructured data
Scientific Visualisation Made Easy
Make any agent harness multimodal-native
Document Image Parsing via Heterogeneous Anchor Prompting”
Python SDK for the Computer Use model Lux, developed by OpenAGI
Large-language-model & vision-language-model based on Linear Attention
AI-powered tool to quickly remove watermarks from images flawlessly
Multi-user UI for managing and running Stable Diffusion workflows tool
A Customizable Image-to-Video Model based on HunyuanVideo
An unsupervised and free tool for image and video dataset analysis
AI Suite for upscaling, interpolating & restoring images/videos
OpenMMLab Model Deployment Framework
computer vision projects | Fun AI projects related to computer vision
AI powered image classification for nudity and documents / id-cards
Real-time music generation using stable diffusion techniques AI
A Strong and Easy-to-use Single View 3D Hand+Body Pose Estimator
Generate text images for training deep learning ocr model