Make your agents learn from experience
Power CLI and Workflow manager for LLMs (core package)
A Simple and Universal Swarm Intelligence Engine
A high-throughput and memory-efficient inference and serving engine
SQL-Driven RAG Engine
LLM inference in C/C++
A tension reasoning engine over 131 S-class problems
Universal LLM Deployment Engine with ML Compilation
Alibaba's high-performance LLM inference engine for diverse apps
Jlama is a modern LLM inference engine for Java
A high-performance inference engine for AI models
A Next-Generation Training Engine Built for Ultra-Large MoE Models
Fast, flexible LLM inference
AI-Powered Data Processing: Use LOTUS to process all of your datasets
A lightweight vLLM implementation built from scratch
950 line, minimal, extensible LLM inference engine built from scratch
AI search engine - self-host with local or cloud LLMs
Mooncake is the serving platform for Kimi
TokenSpeed is a speed-of-light LLM inference engine
Fast Multimodal LLM on Mobile Devices
Query anything (GitHub, Notion, +40 more) with SQL and let LLMs
local-first semantic code search engine
A modular graph-based Retrieval-Augmented Generation (RAG) system
A @ClickHouse fork that supports high-performance vector search
High-performance inference framework for large language models