Port of Facebook's LLaMA model in C/C++
WebAssembly binding for llama.cpp - Enabling on-browser LLM inference
Run Local LLMs on Any Device. Open-source
Clippy, now with some AI
Interface for OuteTTS models
VS Code extension for LLM-assisted code/text completion
Run a full local LLM stack with one command using Docker
Terminal-native coding agent powered by local LLMs
Open-source LLM load balancer and serving platform for hosting LLMs
Structured Outputs
Towards Human-Sounding Speech
Self-hosted ChatGPT-like chatbot powered by Llama models locally
Fast uncensored Gemma model optimized for local chat and coding
JetBrains’ 4B parameter code model for completions
Jan-v1-edge: efficient 1.7B reasoning model optimized for edge devices