Code for running inference with the SAM 3D Body Model 3DB
Convert Google Gemini web into OpenAI-compatible API
GLM-4.5: Open-source LLM for intelligent agents by Z.ai
Genome modeling and design across all domains of life
PyTorch code and models for the DINOv2 self-supervised learning
Open image model at the forefront of design
Open-Source Financial Large Language Models
Open-source large language model family from Tencent Hunyuan
Audio foundation model excelling in audio understanding
Revolutionizing Database Interactions with Private LLM Technology
Models for object and human mesh reconstruction
Contexts Optical Compression
A series of math-specific large language models of our Qwen2 series
Qwen2.5-VL is the multimodal large language model series
Official code base for LeWorldModel: Stable End-to-End Joint-Embedding
Recovering the Visual Space from Any Views
Python SDK for Claude Agent
Tool for exploring and debugging transformer model behaviors
A Unified Framework for Text-to-3D and Image-to-3D Generation
Multimodal-Driven Architecture for Customized Video Generation
General-purpose image editing model that delivers high-fidelity
Tencent Hunyuan Multimodal diffusion transformer (MM-DiT) model
Qwen-Image-Layered: Layered Decomposition for Inherent Editablity
GLM-4-Voice | End-to-End Chinese-English Conversational Model
Mixture-of-Experts Vision-Language Models for Advanced Multimodal