Visual Causal Flow
A Customizable Image-to-Video Model based on HunyuanVideo
Programmatic access to the AlphaGenome model
GLM-4 series: Open Multilingual Multimodal Chat LMs
GLM-4.5: Open-source LLM for intelligent agents by Z.ai
scikit-learn compatible tabular foundation model
Uncommon Objects in 3D dataset
Agentic, Reasoning, and Coding (ARC) foundation models
State-of-the-art (SoTA) text-to-video pre-trained model
Pretrained time-series foundation model developed by Google Research
Inference script for Oasis 500M
Global weather forecasting model using graph neural networks and JAX
CogView4, CogView3-Plus and CogView3(ECCV 2024)
Implementation of the Surya Foundation Model for Heliophysics
Video Object and Interaction Deletion
Qwen3-ASR is an open-source series of ASR models
A Pragmatic VLA Foundation Model
CLIP, Predict the most relevant text snippet given an image
Personalize Any Characters with a Scalable Diffusion Transformer
AI cognitive-enhancement Skills based on Anthropic's J-space
Open-source image generative foundation model
Infinite Worlds with Versatile Interactions
Robust Speech Recognition Across Languages, Dialects
The official PyTorch implementation of Google's Gemma models
Sharp Monocular Metric Depth in Less Than a Second