Low-latency AI inference engine optimized for mobile devices
Context-aware desktop AI assistant that understands screen content
CogView4, CogView3-Plus and CogView3(ECCV 2024)
Open source AI VTuber platform with voice chat and Live2D avatars
Automate native Android apps with AI using accessibility APIs
Qwen3 is the large language model series developed by Qwen team
Repo of Qwen2-Audio chat & pretrained large audio language model
The right way to check the weather
Concatenate a directory full of files into a single prompt
Multilingual sentence & image embeddings with BERT
A python tool that uses GPT-4, FFmpeg, and OpenCV
Adding guardrails to large language models
Dealing with all unstructured data, such as reverse image search
Pushing the Frontier of Long Audio-Visual Generation
[CVPR 2026 Oral] VGGT Omega
Flexible Photo Recrafting While Preserving Your Identity
Bailing is a voice dialogue robot similar to GPT-4o
Build Vision Agents quickly with any model or video provider
Chinese and English multimodal conversational language model
Tensor search for humans
Deep Understanding AI Agents
Framework for building, orchestrating, and deploying AI agents
Offical Implementation for "Recursive Multi-Agent Systems"
Search all of YouTube from the command line
Qwen3-ASR is an open-source series of ASR models