HY-Motion model for 3D character animation generation
A GPT-4o Level MLLM for Vision, Speech and Multimodal Live Streaming
A Systematic Framework for Interactive World Modeling
DeepSeek Coder: Let the Code Write Itself
Run GLM-5.2 (744B MoE) on a 25GB-RAM consumer machine
Diversity-driven optimization and large-model reasoning ability
Chinese and English multimodal conversational language model
super expressive prompting model based on ltx2.3
Capable of understanding text, audio, vision, video
GLM-4.5: Open-source LLM for intelligent agents by Z.ai
GLM-5: From Vibe Coding to Agentic Engineering
Advancing Open-source World Models
Global weather forecasting model using graph neural networks and JAX
code for Mesh R-CNN, ICCV 2019
Generating Immersive, Explorable, and Interactive 3D Worlds
CodeGeeX: An Open Multilingual Code Generation Model (KDD 2023)
Qwen3-omni is a natively end-to-end, omni-modal LLM
An easy 1-click way to create beautiful artwork on your PC using AI
Open-source framework for intelligent speech interaction
Large Multimodal Models for Video Understanding and Editing
Bidirectional token-classification model for identifiable info
Genome modeling and design across all domains of life
Achieving 3+ generation speedup on reasoning tasks
Ultra-Efficient LLMs on End Device
Pretrained time-series foundation model developed by Google Research