Port of Facebook's LLaMA model in C/C++
Python bindings for llama.cpp
Run Local LLMs on Any Device. Open-source
Interface for OuteTTS models
Personal AI, On Personal Devices
Claude Code, but it runs on your Mac for free
Powerful Android AI agent with tools, automation, and Linux shell
Run a full local LLM stack with one command using Docker
Your Personal AI Assistant; easy to install, deploy on local or coud
Run Bonsai (1-bit) and Ternary-Bonsai language models locally
Structured Outputs
An easy-to-understand framework for LLM samplers
Qwen3 is the large language model series developed by Qwen team
Oobabooga - The definitive Web UI for local AI, with powerful features
A proxy server for multiple ollama instances with Key security
Towards Human-Sounding Speech
First class Sublime Text AI assistant with gpt-5, Opus 4.6, Gemini 3
Performance-optimized AI inference on your GPUs
GLM-4 series: Open Multilingual Multimodal Chat LMs
Inference Llama 2 in one file of pure C
Run GGUF models easily with a UI or API. One File. Zero Install.
FreeAskInternet is a completely free running search aggregator
Run any Llama 2 locally with gradio UI on GPU or CPU from anywhere
Chinese LLaMA & Alpaca large language model + local CPU/GPU training