Port of Facebook's LLaMA model in C/C++
WebAssembly binding for llama.cpp - Enabling on-browser LLM inference
Run Local LLMs on Any Device. Open-source
Clippy, now with some AI
VS Code extension for LLM-assisted code/text completion
Open-source LLM load balancer and serving platform for hosting LLMs