SOTA discrete acoustic codec models with 40/75 tokens per second
48khz stereo neural audio codec for general audio
Implementation of NÜWA, attention network for text to video synthesis
Implementation of RQ Transformer, autoregressive image generation
Dynamic Generalized Relevance Learning Vector Quantization
dimensionality-recursive vector quantization