Fast inference engine for Transformer models
Miso TTS is an 8 billion, highly emotive text-to-speech model
A HEVC/H.265 Web Player
A Foundation Model for the Language of Financial Markets
MiniMax H3 inference engine for Mac computers
A simple but complete full-attention transformer
Data manipulation and transformation for audio signal processing
Accurate × Fast × Comprehensive
Go library for the TOML file format
Segmentation models with pretrained backbones. PyTorch
Pretrained time-series foundation model developed by Google Research
Wrangling Untrusted File Formats Safely
Open-source industrial-grade ASR models
Reliable, open-source crash reporting for iOS, macOS and tvOS
Multimodal model achieving SOTA performance
TorchMultimodal is a PyTorch library
End-to-end speech processing toolkit
LLM training code for MosaicML foundation models
Boilerplate-free Kotlin config library for loading configuration files
OpenAI swift async text to image for SwiftUI app using OpenAI
AV1 Image File Format Specification - ISO-BMFF/HEIF derivative
Small script to encode to H.264/AVC video
A tool for transcoding lossless audio files
Extremely fast compression algorithm