Document (PDF, Word, PPTX ...) extraction and parse API
Parse files for optimal RAG
A high-quality PDF to Markdown tool based on large language model
Contexts Optical Compression
Open image model at the forefront of design
Generate blog articles from video or audio
AI framework for automated short video creation and editing tools
Reading book source
Autonomous LLM agent for end-to-end data science workflows
OCR expert VLM powered by Hunyuan's native multimodal architecture
Renderer for the harmony response format to be used with gpt-oss
The official implementation of RAPTOR
Structured data extraction and instruction calling with ML, LLM
Voice Recognition to Text Tool
Using AI models to automatically provide commentary and edit videos
Knowledge Graph Generation from Any Text
Public opinion analysis system
OCR model for complex documents with layout-aware structured outputs
Generate audiobooks from e-books
A Family of Open Sourced Music Foundation Models
AI agent to evaluate and score resumes
Document content and metadata extraction microservice
A simple, high-quality voice conversion tool focused on ease of use
AI-Researcher: Autonomous Scientific Innovation
A modular graph-based Retrieval-Augmented Generation (RAG) system