From Images to High-Fidelity 3D Assets
gpt-oss-120b and gpt-oss-20b are two open-weight language models
Python inference and LoRA trainer package for the LTX-2 audio–video
Moonshot's most powerful AI model
Wan2.2: Open and Advanced Large-Scale Video Generative Model
Long-form streaming TTS system for multi-speaker dialogue generation
Clean and efficient FP8 GEMM kernels with fine-grained scaling
Repo for SeedVR2 & SeedVR
Qwen3-VL, the multimodal large language model series by Alibaba Cloud
Image generation model with single-stream diffusion transformer
Official inference repo for FLUX.2 models
Diversity-driven optimization and large-model reasoning ability
Pretrained time-series foundation model developed by Google Research
The official PyTorch implementation of Google's Gemma models
Instructions on how to use the Realtime API on Microcontrollers
State of the art LLM and coding model
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
LLM-based Reinforcement Learning audio edit model
OpenAI’s compact 20B open model for fast, agentic, and local use
OpenAI’s open-weight 120B model optimized for reasoning and tooling