Fast and Universal 3D reconstruction model for versatile tasks
Open-source multi-speaker long-form text-to-speech model
Audio foundation model excelling in audio understanding
Qwen3-ASR is an open-source series of ASR models
Research code artifacts for Code World Model (CWM)
Language modeling in a sentence representation space
Multi-modal large language model designed for audio understanding
Encoder of greater-than-word length text trained on a variety of data
Code for the paper Hybrid Spectrogram and Waveform Source Separation
The official pytorch implementation of our paper