Run GLM-5.2 (744B MoE) on a 25GB-RAM consumer machine
Port of OpenAI's Whisper model in C/C++
Oobabooga - The definitive Web UI for local AI, with powerful features
A community-supported supercharged version of paperless
A 2.78-trillion-parameter Kimi K3 running inference on a single CPU
Running large language models on a single GPU
Supercharge Your LLM with the Fastest KV Cache Layer
Scalable, fast, and disk-friendly vector search in Postgres
Your best AI pair programmer in VS Code
A system monitoring tool that exposes system metrics
Find the local LLM that actually runs and performs best
Context data platform for building observable, self-learning AI agents
Run the full 2.78-trillion-parameter Kimi K3 model
Local RAG engine for private multimodal knowledge search on devices
The Cloud-Native API Gateway
A Powerful Desktop Full-Text Search Engine, Just Like Local Google.
StudioOllamaUI is a local, portable interface for Ollama
Lightweight inference library for ONNX files, written in C++
Ray Aviary - evaluate multiple LLMs easily
Renren Film and Television bot, fully connected to Renren resources
Virtual Assistant Maintenance System
A Python/Pytorch app for easily synthesising human voices
A graphical frontend to tesseract-ocr
A supercharged version of paperless, scan, index and archive docs
A distributed cross-platform Telegram Bot