Port of OpenAI's Whisper model in C/C++
Connect home devices into a powerful cluster to accelerate LLM
High-performance neural network inference framework for mobile
The Triton Inference Server provides an optimized cloud
Easy-to-use deep learning framework with 3 key features
Lightweight anchor-free object detection model
Deep learning inference framework optimized for mobile platforms