Make your agents learn from experience
Power CLI and Workflow manager for LLMs (core package)
A Simple and Universal Swarm Intelligence Engine
Alibaba's high-performance LLM inference engine for diverse apps
Fast, flexible LLM inference
Mooncake is the serving platform for Kimi
Jlama is a modern LLM inference engine for Java
AI-Powered Data Processing: Use LOTUS to process all of your datasets
Query anything (GitHub, Notion, +40 more) with SQL and let LLMs
A high-throughput and memory-efficient inference and serving engine
SQL-Driven RAG Engine
All-in-one AI companion! Desktop girlfriend + virtual streamer
A Next-Generation Training Engine Built for Ultra-Large MoE Models
Universal LLM Deployment Engine with ML Compilation
local-first semantic code search engine
High-performance inference framework for large language models
A modular graph-based Retrieval-Augmented Generation (RAG) system
A tension reasoning engine over 131 S-class problems
A high-performance inference engine for AI models
TokenSpeed is a speed-of-light LLM inference engine
Run AI models locally on your machine with node.js bindings for llama
A lightweight vLLM implementation built from scratch
OpenAI API client for Kotlin with multiplatform capabilities
LLM inference in C/C++
AI search engine - self-host with local or cloud LLMs