Run models like Kimi-K2.5, GLM-5, DeepSeek, gpt-oss, Gemma, Qwen etc.
OmniRoute is an AI gateway for multi-provider LLM
Port of Facebook's LLaMA model in C/C++
Advanced LLM-powered brute-force tool combining AI intelligence
LLM Frontend for Power Users
Official code repo for the O'Reilly Book
The all-in-one Desktop & Docker AI application with full RAG and AI
CLI proxy that reduces LLM token consumption
AI Coding agent for the terminal
Open-source, high-performance AI model with advanced reasoning
Run Local LLMs on Any Device. Open-source
A high-throughput and memory-efficient inference and serving engine
Distribute and run LLMs with a single file
Powerful AI language model (MoE) optimized for efficiency/performance
From Vibe Coding to Agentic Engineering
The media player for language learning, with dual subtitles
Drag & drop UI to build your customized LLM flow
Use any SDK to call 100+ LLMs
Fully automatic censorship removal for language models
OpenAI-compatible proxy that aggregates free-tier keys from ~14 AI
Project aimed at extracting, exporting, and analyzing chat records
Python bindings for llama.cpp
Clippy, now with some AI
157 models, 30 providers, one command to find what runs on hardware
lightweight package to simplify LLM API calls