PyTorch extensions for fast R&D prototyping and Kaggle farming
Robust Speech Recognition via Large-Scale Weak Supervision
A SOTA open-source image editing model
A Family of Open Foundation Models for Code Intelligence
Miso TTS is an 8 billion, highly emotive text-to-speech model
Fast inference engine for Transformer models
Industrial-level controllable zero-shot text-to-speech system
Pretrained time-series foundation model developed by Google Research
Provides code for running inference with the SegmentAnything Model
MiniMax H3 inference engine for Mac computers
Data manipulation and transformation for audio signal processing
End-to-end speech processing toolkit
Accurate × Fast × Comprehensive
A simple but complete full-attention transformer
Open-source industrial-grade ASR models
Multimodal model achieving SOTA performance
LLM training code for MosaicML foundation models
OpenAI swift async text to image for SwiftUI app using OpenAI
A MATLAB package for modelling multivariate stimulus-response data
A Conversational Speech Generation Model
The unofficial python package that returns response of Google Bard
DeepSeek LLM: Let there be answers
Implementation of DALL-E 2, OpenAI's updated text-to-image synthesis
Consistency Distilled Diff VAE
Basaran, an open-source alternative to the OpenAI text completion API