Convert Google Gemini web into OpenAI-compatible API
PyTorch code and models for the DINOv2 self-supervised learning
Open image model at the forefront of design
GLM-4.5: Open-source LLM for intelligent agents by Z.ai
Open-Source Financial Large Language Models
Open-source large language model family from Tencent Hunyuan
Audio foundation model excelling in audio understanding
Revolutionizing Database Interactions with Private LLM Technology
Models for object and human mesh reconstruction
Contexts Optical Compression
A series of math-specific large language models of our Qwen2 series
Qwen2.5-VL is the multimodal large language model series
Recovering the Visual Space from Any Views
Python SDK for Claude Agent
Tool for exploring and debugging transformer model behaviors
A Unified Framework for Text-to-3D and Image-to-3D Generation
Multimodal-Driven Architecture for Customized Video Generation
General-purpose image editing model that delivers high-fidelity
Tencent Hunyuan Multimodal diffusion transformer (MM-DiT) model
Qwen-Image-Layered: Layered Decomposition for Inherent Editablity
GLM-4-Voice | End-to-End Chinese-English Conversational Model
Mixture-of-Experts Vision-Language Models for Advanced Multimodal
DeepSeek Coder: Let the Code Write Itself
Generating Immersive, Explorable, and Interactive 3D Worlds
Community plugin marketplace for Claude Cowork and Claude Code