SQL-Driven RAG Engine
Universal LLM Deployment Engine with ML Compilation
950 line, minimal, extensible LLM inference engine built from scratch
Run a 1-billion parameter LLM on a $10 board with 256MB RAM
Fast Multimodal LLM on Mobile Devices
Emscripten: An LLVM-to-WebAssembly Compiler
Llama 2 Everywhere (L2E)