A Family of Open Sourced Music Foundation Models
Qwen-Image is a powerful image generation foundation model
Industrial-level controllable zero-shot text-to-speech system
gpt-oss-120b and gpt-oss-20b are two open-weight language models
Programmatic access to the AlphaGenome model
Official repository for LTX-Video
Accurate × Fast × Comprehensive
Qwen3-TTS is an open-source series of TTS models
Text and image to video generation: CogVideoX and CogVideo
AlphaFold 3 inference pipeline
Convert Google Gemini web into OpenAI-compatible API
Models for object and human mesh reconstruction
Open-source multi-speaker long-form text-to-speech model
Visual Causal Flow
High-Resolution Image Synthesis with Latent Diffusion Models
State-of-the-art TTS model under 25MB
MapAnything: Universal Feed-Forward Metric 3D Reconstruction
Generating Immersive, Explorable, and Interactive 3D Worlds
Qwen3 is the large language model series developed by Qwen team
Recovering the Visual Space from Any Views
Hackable and optimized Transformers building blocks
Revolutionizing Database Interactions with Private LLM Technology
Qwen3-Coder is the code version of Qwen3
Qwen2.5-VL is the multimodal large language model series
Foundation Models for Time Series