Port of Facebook's LLaMA model in C/C++
LLM inference in C/C++
Tensor library for machine learning
lightweight, standalone C++ inference engine for Google's Gemma models
Low-latency machine code generation
Production ready toolkit to run AI locally
A lightweight, lightning-fast, in-process vector database
Serving system for machine learning models
Fast Multimodal LLM on Mobile Devices
mods to the Festival sokoban solver to run on OSX + Win + linux
Open deep learning compiler stack for cpu, gpu
The easiest C++ way to deal with constraints !