Run Local LLMs on Any Device. Open-source
ONNX Runtime: cross-platform, high performance ML inferencing
TT-NN operator library, and TT-Metalium low level kernel programming
C++ library for high performance inference on NVIDIA GPUs
High-performance neural network inference framework for mobile
MNN is a blazing fast, lightweight deep learning framework
A GPU-accelerated library containing highly optimized building blocks
Set of comprehensive computer vision & machine intelligence libraries
QodeAssist is an AI-powered coding assistant plugin for Qt Creator
Modern, Header-only C++ bindings for the Ollama API
Run GGUF models easily with a UI or API. One File. Zero Install.
Easy-to-use deep learning framework with 3 key features
Guide to deploying deep-learning inference networks
Implements a reference architecture for creating information systems
Deep learning inference framework optimized for mobile platforms
Uniform deep learning inference framework for mobile
Embeddable scripting runtime for live behavior, AI, and automation.