Port of Facebook's LLaMA model in C/C++
Python bindings for llama.cpp
Run Local LLMs on Any Device. Open-source
Interface for OuteTTS models
Personal AI, On Personal Devices
Claude Code, but it runs on your Mac for free
Powerful Android AI agent with tools, automation, and Linux shell
Your Personal AI Assistant; easy to install, deploy on local or coud
Run Bonsai (1-bit) and Ternary-Bonsai language models locally
Run a full local LLM stack with one command using Docker
Structured Outputs
Qwen3 is the large language model series developed by Qwen team
Oobabooga - The definitive Web UI for local AI, with powerful features
Towards Human-Sounding Speech
GLM-4 series: Open Multilingual Multimodal Chat LMs
Performance-optimized AI inference on your GPUs
Inference Llama 2 in one file of pure C
Run GGUF models easily with a UI or API. One File. Zero Install.
Run any Llama 2 locally with gradio UI on GPU or CPU from anywhere
Chinese LLaMA & Alpaca large language model + local CPU/GPU training
JetBrains’ 4B parameter code model for completions