Qwen-Image is a powerful image generation foundation model
All-in-one WebUI for AI generative image and video creation
Guiding Instruction-based Image Editing via Multimodal Large Language
CogView4, CogView3-Plus and CogView3(ECCV 2024)
Flutter-based cross-platform app integrating major AI models
The free, Open Source alternative to OpenAI, Claude and others
Tensor search for humans
Gracefully face hCaptcha challenge with multimodal llms
Capable of understanding text, audio, vision, video
An LLM-based presentation generation platform
AI-powered code assistant for Vim. OpenAI and ChatGPT plugin for Vim
Multilingual sentence & image embeddings with BERT
Full stack framework for building cross-platform mobile AI apps
Phi-3.5 for Mac: Locally-run Vision and Language Models
A powerful tool for creating datasets for LLM fine-tuning
InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System
Moonshot's most powerful AI model
Production-ready AI chat. Start here and make it your own
Qwen3-omni is a natively end-to-end, omni-modal LLM
The Multi-Agent Framework
Open source libraries and APIs to build custom preprocessing pipelines
Fast and efficient unstructured data extraction
Open-source evaluation toolkit of large multi-modality models (LMMs)
Fast Multimodal LLM on Mobile Devices
The open source codebase powering HuggingChat