Open-source image generative foundation model
4M: Massively Multimodal Masked Modeling
Claude Code image, a one-stop open source transit service
Accurate × Fast × Comprehensive
OCR expert VLM powered by Hunyuan's native multimodal architecture
DeepSeek Harness (DSH) Web
Repo for SeedVR2 & SeedVR
Large Multimodal Models for Video Understanding and Editing
Implementation of the Surya Foundation Model for Heliophysics
Collection of Gemma 3 variants that are trained for performance
Pretrained time-series foundation model developed by Google Research
Proxy that exposes Antigravity provided claude / gemini models
New set of lightweight state-of-the-art, open foundation models
tiktoken is a fast BPE tokeniser for use with OpenAI's models
Repo of Qwen2-Audio chat & pretrained large audio language model
MiniMax-M2, a model built for Max coding & agentic workflows
A Powerful Native Multimodal Model for Image Generation
State of the art LLM and coding model
Official implementation of DreamCraft3D
Global weather forecasting model using graph neural networks and JAX
Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence
The official PyTorch implementation of Google's Gemma models
A 0.1B Omni model trained from scratch
Block Diffusion for Ultra-Fast Speculative Decoding
MOSS‑TTS Family open‑source speech and sound generation model