Official inference repo for FLUX.2 models
A Customizable Image-to-Video Model based on HunyuanVideo
High-Resolution Image Synthesis with Latent Diffusion Models
The official Meta Llama 3 GitHub site
High-Quality Voice Cloning TTS for 600+ Languages
A Family of Open Sourced Music Foundation Models
Instant voice cloning by MIT and MyShell. Audio foundation model
Rebuild the object in a reference image as a code-only, procedural
AI generative media user experience highlighting use of APIs
Utilities intended for use with Llama models
A sound cloning tool with a web interface, using your voice
An open-source toolkit for BigMac-style pipeline-parallel training
Tencent Hunyuan Multimodal diffusion transformer (MM-DiT) model
Clone a voice in 5 seconds to generate arbitrary speech in real-time
Codex plugin that turns attached object images into code-only
Suite of reference architectures for building GPU-accelerated vision
Tokenizer-Free TTS for Multilingual Speech Generation
Generative AI reference workflows
This repository contains code released by Google Research
Personalize Any Characters with a Scalable Diffusion Transformer
150+ quantitative finance Python programs
A high-quality rapid TTS voice cloning model
TTS for Context-Aware Speech Generation and True-to-Life Voice Cloning
Interface for OuteTTS models
Specification and documentation for Agent Skills