Run models like Kimi-K2.5, GLM-5, DeepSeek, gpt-oss, Gemma, Qwen etc.
Port of Facebook's LLaMA model in C/C++
High-performance code intelligence MCP server
ESP32 desk dashboard that shows Claude Code usage
A 2.78-trillion-parameter Kimi K3 running inference on a single CPU
Run the full 2.78-trillion-parameter Kimi K3 model
Run GLM-5.2 (744B MoE) on a 25GB-RAM consumer machine
AI video generator optimized for low VRAM and older GPUs use
Low-latency AI inference engine optimized for mobile devices
Flux 2 image generation model pure C inference
DeepSeek 4 Flash local inference engine for Metal
llama and other large language models on iOS and MacOS offline
AI macOS app for real-time coding interview coaching assistance
Locally run an Instruction-Tuned Chat-Style LLM
mujoco-py allows using MuJoCo from Python 3
Multiagent simulator of road traffic in Qt/C++ and OpenStreetMap.
Library written in C with Python API for IPv6 networking