CLI proxy that reduces LLM token consumption
Fast, flexible LLM inference
157 models, 30 providers, one command to find what runs on hardware
Fast, local-first web content extraction for LLMs
Beautiful git diff viewer, generate commits with AI
Distributed LLM and StableDiffusion inference
A high-performance inference engine for AI models
An ecosystem of Rust libraries for working with large language models
A program that provides LLMs with ability to complete complex tasks