Diversity-driven optimization and large-model reasoning ability
Advancing Open-source World Models
GLM-4.5: Open-source LLM for intelligent agents by Z.ai
Official code base for LeWorldModel: Stable End-to-End Joint-Embedding
An experimental version of DeepSeek model
Infinite Worlds with Versatile Interactions
Programmatic access to the AlphaGenome model
Visual Causal Flow
Codex plugin that turns attached object images into code-only
Ling is a MoE LLM provided and open-sourced by InclusionAI
CLIP, Predict the most relevant text snippet given an image
Repo for SeedVR2 & SeedVR
Recovering the Visual Space from Any Views
4M: Massively Multimodal Masked Modeling
OCR expert VLM powered by Hunyuan's native multimodal architecture
A SOTA open-source image editing model
High-Fidelity and Controllable Generation of Textured 3D Assets
Large Multimodal Models for Video Understanding and Editing
Open-source image generative foundation model
Collection of Gemma 3 variants that are trained for performance
Block Diffusion for Ultra-Fast Speculative Decoding
Accurate × Fast × Comprehensive
Pretrained time-series foundation model developed by Google Research
code for Mesh R-CNN, ICCV 2019
Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence