AWS IoT FleetWise Edge Agent
On-device AI across mobile, embedded and edge for PyTorch
LiteRT-LM is Google's production-ready inference framework
A lightweight, lightning-fast, in-process vector database
Diffusion model(SD,Flux,Wan,Qwen Image,Z-Image,...) inference
OpenVINO™ Toolkit repository
LiteRT, successor to TensorFlow Lite
Infrastructure to enable deployment of ML models
Fast Multimodal LLM on Mobile Devices
An Easy-to-Use and High-Performance AI Deployment Framework
A lightweight 3D Morphable Face Model library in modern C++
MNN is a blazing fast, lightweight deep learning framework
CUDA Templates for Linear Algebra Subroutines
A retargetable MLIR-based machine learning compiler runtime toolkit
A scalable inference server for models optimized with OpenVINO
Easy-to-use Speech Toolkit including Self-Supervised Learning model
Cross-platform, customizable ML solutions
Fast, Sharp & Reliable Agentic Intelligence
mlpack: a scalable C++ machine learning library
Clean and efficient FP8 GEMM kernels with fine-grained scaling
Testing tool for modeling GUI transitions
C++ implementation of ChatGLM-6B & ChatGLM2-6B & ChatGLM3 & GLM4(V)
Free Streaming Bot: Compatible with Twitch, YouTube and Facebook
Runtime extension of Proximus enabling Deployment on AMD Ryzen™ AI
Award-winning modern data processing SDK in C++20