High-performance library for gradient boosting on decision trees
High-speed Large Language Model Serving for Local Deployment
Gradient boosting framework based on decision tree algorithms
MNN is a blazing fast, lightweight deep learning framework
FlashMLA: Efficient Multi-head Latent Attention Kernels
Mooncake is the serving platform for Kimi
fast C++ library for GPU linear algebra & scientific computing
Type in any Windows app at the speed of speech.
Easy-to-use deep learning framework with 3 key features
Lightweight, Portable, Flexible Distributed/Mobile Deep Learning
Deep learning inference framework optimized for mobile platforms
Caffe, a fast open framework for deep learning
A fast open framework for deep learning