Official inference repo for FLUX.2 models
A Customizable Image-to-Video Model based on HunyuanVideo
Infinite Canvas Workbench for AI creation integrates AI generation
High-Resolution Image Synthesis with Latent Diffusion Models
The official Meta Llama 3 GitHub site
High-Quality Voice Cloning TTS for 600+ Languages
Rebuild the object in a reference image as a code-only, procedural
Instant voice cloning by MIT and MyShell. Audio foundation model
AI generative media user experience highlighting use of APIs
From Beginner to Master · Orange Book Series
This repository contains code released by Google Research
A Family of Open Sourced Music Foundation Models
super expressive prompting model based on ltx2.3
Tokenizer-Free TTS for Multilingual Speech Generation
Codex plugin that turns attached object images into code-only
Utilities intended for use with Llama models
Tencent Hunyuan Multimodal diffusion transformer (MM-DiT) model
AI coding jargon, explained in plain English
Personal notes from Wu Enda's machine learning course
Clone a voice in 5 seconds to generate arbitrary speech in real-time
Generative AI reference workflows
An open-source toolkit for BigMac-style pipeline-parallel training
OpenAI gpt-image-2 API
Suite of reference architectures for building GPU-accelerated vision
Token-Oriented Object Notation (TOON)