Alibaba's high-performance LLM inference engine for diverse apps
1B text generation model based on the HRM architecture
Document (PDF, Word, PPTX ...) extraction and parse API
High-performance inference server for text embeddings models API layer
Hypernetworks that adapt LLMs for specific benchmark tasks
TTS with kokoro and onnx runtime
Industrial-level controllable zero-shot text-to-speech system
The media player for language learning, with dual subtitles
Claude Code skill that removes signs of AI-generated writing from text
High-Quality Voice Cloning TTS for 600+ Languages
Qwen3-TTS is an open-source series of TTS models
Code for openai.fm, a demo for the OpenAI Speech API
Tokenizer-Free TTS for Multilingual Speech Generation
MiniMax H3 is a general-purpose, omni-modal generative system
A TTS that fits in your CPU (and pocket)
A Powerful Native Multimodal Model for Image Generation
Contexts Optical Compression
Official inference repo for FLUX.1 models
Robust Speech Recognition via Large-Scale Weak Supervision
super expressive prompting model based on ltx2.3
Code for running inference and finetuning with SAM 3 model
JavaScript OCR and text extraction for images and PDFs
Use Microsoft Edge's online text-to-speech service from Python
Lightning-fast, on-device TTS, running natively via ONNX
A typeface that protects written content by poisoning unauthorized AI