Python inference and LoRA trainer package for the LTX-2 audio–video
A Powerful Native Multimodal Model for Image Generation
A 0.1B Omni model trained from scratch
OCR expert VLM powered by Hunyuan's native multimodal architecture
Native and Compact Structured Latents for 3D Generation
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
Official Python inference and LoRA trainer package
gpt-oss-120b and gpt-oss-20b are two open-weight language models
Official inference repo for FLUX.2 models
Convert Google Gemini web into OpenAI-compatible API
New family of code large language models (LLMs)
GLM-4 series: Open Multilingual Multimodal Chat LMs
Open-weight, large-scale hybrid-attention reasoning model
Qwen3-Coder is the code version of Qwen3
Open image model at the forefront of design
Qwen-Image is a powerful image generation foundation model
FAIR Sequence Modeling Toolkit 2
High-Fidelity and Controllable Generation of Textured 3D Assets
Open Multilingual Multimodal Chat LMs
OpenAI’s compact 20B open model for fast, agentic, and local use