From Images to High-Fidelity 3D Assets
Moonshot's most powerful AI model
Python inference and LoRA trainer package for the LTX-2 audio–video
gpt-oss-120b and gpt-oss-20b are two open-weight language models
Long-form streaming TTS system for multi-speaker dialogue generation
Wan2.2: Open and Advanced Large-Scale Video Generative Model
Official inference repo for FLUX.2 models
Designed for text embedding and ranking tasks
Image generation model with single-stream diffusion transformer
Diversity-driven optimization and large-model reasoning ability
The official PyTorch implementation of Google's Gemma models
Qwen3-VL, the multimodal large language model series by Alibaba Cloud
Clean and efficient FP8 GEMM kernels with fine-grained scaling
State of the art LLM and coding model
Instructions on how to use the Realtime API on Microcontrollers
Pretrained time-series foundation model developed by Google Research
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
Towards Robust Blind Face Restoration with Codebook Lookup Transformer
OpenAI’s compact 20B open model for fast, agentic, and local use
OpenAI’s open-weight 120B model optimized for reasoning and tooling