From Images to High-Fidelity 3D Assets
Convert Google Gemini web into OpenAI-compatible API
Phi-3.5 for Mac: Locally-run Vision and Language Models
26m function call model that runs on incredibly small devices
Open-source image generative foundation model
High-Resolution Image Synthesis with Latent Diffusion Models
Personalize Any Characters with a Scalable Diffusion Transformer
State-of-the-art TTS model under 25MB
Qwen-Image is a powerful image generation foundation model
An Efficient Agentic Model for Computer Use
Multimodal-Driven Architecture for Customized Video Generation
Code for running inference and finetuning with SAM 3 model
A Pragmatic VLA Foundation Model
Diffusion model(SD,Flux,Wan,Qwen Image,Z-Image,...) inference
Collection of Gemma 3 variants that are trained for performance
Lets make video diffusion practical
A Powerful Native Multimodal Model for Image Generation
Robust Speech Recognition Across Languages, Dialects
Open image model at the forefront of design
Qwen3-TTS is an open-source series of TTS models
HY-Motion model for 3D character animation generation
Powerful AI language model (MoE) optimized for efficiency/performance
Pokee Deep Research Model Open Source Repo
Diffusion Bee is the easiest way to run Stable Diffusion locally
The official repo of Qwen chat & pretrained large language model