Two-phase DeepSeek Harness preset
950 line, minimal, extensible LLM inference engine built from scratch
Port of Facebook's LLaMA model in C/C++
LLM inference in C/C++
Apple Intelligence from the command line
A minimal LLM chat app that runs entirely in your browser
CLI proxy that reduces LLM token consumption
Structured outputs for llms
Run a 1-billion parameter LLM on a $10 board with 256MB RAM
A lightweight vLLM implementation built from scratch
LangChain4j is an open-source Java library
The PHP Agentic Framework to build production-ready AI driven apps
All-in-one WebUI for AI generative image and video creation
Production ready toolkit to run AI locally
Implement a ChatGPT-like LLM in PyTorch from scratch, step by step
Open-source evaluation toolkit of large multi-modality models (LMMs)
Deploy your agentic worfklows to production
Create architecture diagrams from code automatically using LLMs
The most powerful MCP Slack Server with no permission requirements
From Paper to Presentation in One Click
Personal AI Notebooks. Organize files & webpages and generate notes
LLM training in simple, raw C/CUDA
A high-performance inference engine for AI models
LightLLM is a Python-based LLM (Large Language Model) inference
Build a modern LLM from scratch. Every line commented