A SOTA open-source image editing model
A Family of Open Foundation Models for Code Intelligence
Miso TTS is an 8 billion, highly emotive text-to-speech model
Industrial-level controllable zero-shot text-to-speech system
Pretrained time-series foundation model developed by Google Research
MiniMax H3 inference engine for Mac computers
Accurate × Fast × Comprehensive
Open-source industrial-grade ASR models
Multimodal model achieving SOTA performance
A Conversational Speech Generation Model
DeepSeek LLM: Let there be answers
Open-source pre-training implementation of Google's LaMDA in PyTorch
Code release for "Masked-attention Mask Transformer
PyTorch implementation of MAE
Facebook AI Research Sequence-to-Sequence Toolkit
Real Time Speech Enhancement in the Waveform Domain (Interspeech 2020)
Efficient 309B omnimodal MoE for coding, agents, vision, and audio
1T omnimodal MoE model for coding, agents, and long-horizon reasoning
Flexible text-to-text transformer model for multilingual NLP tasks
Summarization model fine-tuned on CNN/DailyMail articles
Efficient multimodal MoE model for coding, tools, and reasoning