Open source libraries and APIs to build custom preprocessing pipelines
Document (PDF, Word, PPTX ...) extraction and parse API
Uses Qwen3-ASR, local LLM, Whisper, TEN-VAD
An LLM-powered knowledge curation system that researches topics
AI-Powered Data Processing: Use LOTUS to process all of your datasets
Python bindings for llama.cpp
All-in-one WebUI for AI generative image and video creation
Structured data extraction and instruction calling with ML, LLM
A system for agentic LLM-powered data processing and ETL
A high-quality PDF to Markdown tool based on large language model
One-stop solution for creating your digital avatar from chat history
File Parser optimised for LLM Ingestion with no loss
Toolkit for conversational AI
Swirl queries any number of data sources with APIs
LLM based data scientist, AI native data application
AI-powered tool for efficient abstract and PDF screening
SDG is a specialized framework
Enhances Tesseract OCR output using LLMs (local or API)
SQL-Driven RAG Engine
An AI-powered file management tool that ensures privacy
Ongoing research training transformer models at scale
Using AI models to automatically provide commentary and edit videos
Tools for merging pretrained large language models
MoBA: Mixture of Block Attention for Long-Context LLMs
E2M converts various file types (doc, docx, epub, html, htm, url