Qwen3-TTS is an open-source series of TTS models
Achieving 3+ generation speedup on reasoning tasks
Ultra-Efficient LLMs on End Device
MiMo-V2-Flash: Efficient Reasoning, Coding, and Agentic Foundation
Industrial-level controllable zero-shot text-to-speech system
Sharp Monocular Metric Depth in Less Than a Second
Image generation model with single-stream diffusion transformer
PyTorch code and models for the DINOv2 self-supervised learning
Hunyuan Translation Model Version 1.5
Stable Diffusion WebUI Forge is a platform on top of Stable Diffusion
Block Diffusion for Ultra-Fast Speculative Decoding
A Unified Framework for Text-to-3D and Image-to-3D Generation
tiktoken is a fast BPE tokeniser for use with OpenAI's models
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
FlashMLA: Efficient Multi-head Latent Attention Kernels
Powerful open source image generation model
Official DeiT repository
GLM-130B: An Open Bilingual Pre-Trained Model (ICLR 2023)
Large language model developed and released by NVIDIA
Fast uncensored Gemma model optimized for local chat and coding
Speculative-decoding accelerator for the 675B Mistral Large 3
OpenAI’s open-weight 120B model optimized for reasoning and tooling
Compact English sentence embedding model for semantic search tasks
Frontier multimodal MoE model for coding and AI agent workflows
NVFP4 DiffusionGemma model for fast multimodal text generation