HY-Motion model for 3D character animation generation
DeepSeek Coder: Let the Code Write Itself
An experimental version of DeepSeek model
OCR expert VLM powered by Hunyuan's native multimodal architecture
General-purpose image editing model that delivers high-fidelity
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
Tool for exploring and debugging transformer model behaviors
GLM-4.5: Open-source LLM for intelligent agents by Z.ai
Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence
Inference code for scalable emulation of protein equilibrium ensembles
Tongyi Deep Research, the Leading Open-source Deep Research Agent
Qwen-Image is a powerful image generation foundation model
Open Source Speech Language Model
Multimodal Diffusion with Representation Alignment
Stable Diffusion WebUI Forge is a platform on top of Stable Diffusion
MOSS‑TTS Family open‑source speech and sound generation model
ICLR2024 Spotlight: curation/training code, metadata, distribution
Official implementation of DreamCraft3D
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
VGGSfM: Visual Geometry Grounded Deep Structure From Motion
State-of-the-art Image & Video CLIP, Multimodal Large Language Models
Language modeling in a sentence representation space
A SOTA open-source image editing model
ChatGLM-6B: An Open Bilingual Dialogue Language Model
ChatGPT interface with better UI