Make your agents learn from experience
Power CLI and Workflow manager for LLMs (core package)
A Simple and Universal Swarm Intelligence Engine
A high-throughput and memory-efficient inference and serving engine
A tension reasoning engine over 131 S-class problems
Universal LLM Deployment Engine with ML Compilation
AI-Powered Data Processing: Use LOTUS to process all of your datasets
A lightweight vLLM implementation built from scratch
SQL-Driven RAG Engine
A Next-Generation Training Engine Built for Ultra-Large MoE Models
950 line, minimal, extensible LLM inference engine built from scratch
TokenSpeed is a speed-of-light LLM inference engine
local-first semantic code search engine
A modular graph-based Retrieval-Augmented Generation (RAG) system
High-performance inference framework for large language models
Tensor search for humans
Claude + Obsidian knowledge companion
LightLLM is a Python-based LLM (Large Language Model) inference
Parallax is a distributed model serving framework
Build multimodal language agents for fast prototype and production
Request recommended movies, TV shows and anime to Jellyseer/Overseer
GitLab automatic code review tool based on large models
Inference Llama 2 in one file of pure C
Retrieval Augmented Generation (RAG) framework
Ship RAG based LLM web apps in seconds