Image generation model with single-stream diffusion transformer
Production-tested AI infrastructure tools
Tiny vision language model
Inference code for scalable emulation of protein equilibrium ensembles
State of the art LLM and coding model
Proxy that exposes Antigravity provided claude / gemini models
Strong, Economical, and Efficient Mixture-of-Experts Language Model
Lets make video diffusion practical
Open-source image generative foundation model
A Powerful Native Multimodal Model for Image Generation
Advancing Open-source World Models
Official code base for LeWorldModel: Stable End-to-End Joint-Embedding
Diversity-driven optimization and large-model reasoning ability
Moonshot's most powerful AI model
CLIP, Predict the most relevant text snippet given an image
A 0.1B Omni model trained from scratch
Large Multimodal Models for Video Understanding and Editing
MiniMax-M2, a model built for Max coding & agentic workflows
Qwen3-VL, the multimodal large language model series by Alibaba Cloud
Code for running inference with the SAM 3D Body Model 3DB
A Customizable Image-to-Video Model based on HunyuanVideo
Code for running inference and finetuning with SAM 3 model
Infinite Worlds with Versatile Interactions
Qwen3 is the large language model series developed by Qwen team
GLM-4 series: Open Multilingual Multimodal Chat LMs