Document (PDF, Word, PPTX ...) extraction and parse API
Hypernetworks that adapt LLMs for specific benchmark tasks
Qwen-Image is a powerful image generation foundation model
GLM-4-Voice | End-to-End Chinese-English Conversational Model
LLM abstractions that aren't obstructions
A modular graph-based Retrieval-Augmented Generation (RAG) system
Simple, Pythonic building blocks to evaluate LLM applications
Qwen3-omni is a natively end-to-end, omni-modal LLM
Using AI models to automatically provide commentary and edit videos
Unifying 3D Mesh Generation with Language Models
AI-powered code assistant for Vim. OpenAI and ChatGPT plugin for Vim
Knowledge Graph Generation from Any Text
A high-quality PDF to Markdown tool based on large language model
LLM inference server with continuous batching & SSD caching
Toolkit for conversational AI
Build multimodal language agents for fast prototype and production
lightweight package to simplify LLM API calls
Multilingual sentence & image embeddings with BERT
Data Infrastructure providing an approach to multimodal AI workloads
Search all of YouTube from the command line
Retrieval and Retrieval-augmented LLMs
Capable of understanding text, audio, vision, video
Scalable data pre processing and curation toolkit for LLMs
Adding guardrails to large language models
A list of free LLM inference resources accessible via API