FlashInfer: Kernel Library for LLM Serving
A RWKV management and startup tool, full automation, only 8MB
Ready-to-use OCR with 80+ supported languages
C++ library for high performance inference on NVIDIA GPUs
OpenMMLab Model Deployment Framework
Self-contained Machine Learning and Natural Language Processing lib
Lightweight anchor-free object detection model
Deep learning inference framework optimized for mobile platforms
Fast and user-friendly runtime for transformer inference