Apache-2.0 open-source image generation and editing model family
Spark-TTS Inference Code
Multi-lingual large voice generation model, providing inference
High-Resolution Image Synthesis with Latent Diffusion Models
Simple LaTeX parser providing latex-to-unicode and unicode-to-latex
Edit PDF files with Nano Banana
Designed for text embedding and ranking tasks
Agent harness to make your slop code well-engineered and beautiful
Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence
TTS model capable of streaming conversational audio in realtime
A Powerful Native Multimodal Model for Image Generation
Unified web UI for training and running open models locally
HunyuanVideo: A Systematic Framework For Large Video Generation Model
Instant voice cloning by MIT and MyShell. Audio foundation model
Open-source image generative foundation model
Handwritten Text Recognition (HTR) system implemented with TensorFlow
Official inference repo for FLUX.2 models
Offline inference engine for art, real-time voice conversations
Point excavation and submission book
Library for OCR-related tasks powered by Deep Learning
Lightweight Markdown-only skills for autonomous ML research
Sample code and notebooks for Generative AI on Google Cloud
Collection of Gemma 3 variants that are trained for performance
Sublime Text plugin for EditorConfig
The beginning of scalable pixel-native search