Lightweight inference library for ONNX files, written in C++
High quality, fast, modular reference implementation of SSD in PyTorch
Database system for building simpler and faster AI-powered application
Serve machine learning models within a Docker container
Self-contained Machine Learning and Natural Language Processing lib
Implementation of "Tree of Thoughts
llama.go is like llama.cpp in pure Golang
The deep learning toolkit for speech-to-text
Guide to deploying deep-learning inference networks
Training & Implementation of chatbots leveraging GPT-like architecture
Deep learning inference framework optimized for mobile platforms
Uniform deep learning inference framework for mobile
Deploy a ML inference service on a budget in 10 lines of code
Fast and user-friendly runtime for transformer inference