Alibaba's high-performance LLM inference engine for diverse apps
1B text generation model based on the HRM architecture
Document (PDF, Word, PPTX ...) extraction and parse API
High-performance inference server for text embeddings models API layer
Hypernetworks that adapt LLMs for specific benchmark tasks
TTS with kokoro and onnx runtime
Industrial-level controllable zero-shot text-to-speech system
The media player for language learning, with dual subtitles
High-Quality Voice Cloning TTS for 600+ Languages
Claude Code skill that removes signs of AI-generated writing from text
Qwen3-TTS is an open-source series of TTS models
Code for openai.fm, a demo for the OpenAI Speech API
Tokenizer-Free TTS for Multilingual Speech Generation
MiniMax H3 is a general-purpose, omni-modal generative system
A Powerful Native Multimodal Model for Image Generation
Official inference repo for FLUX.1 models
Contexts Optical Compression
Robust Speech Recognition via Large-Scale Weak Supervision
A TTS that fits in your CPU (and pocket)
super expressive prompting model based on ltx2.3
Lightning-fast, on-device TTS, running natively via ONNX
Code for running inference and finetuning with SAM 3 model
Use Microsoft Edge's online text-to-speech service from Python
JavaScript OCR and text extraction for images and PDFs
Open Frontier Intelligence