Fast stable diffusion on CPU and AI PC
High-speed Large Language Model Serving for Local Deployment
Fast inference engine for Transformer models
Faster Whisper transcription with CTranslate2
High-performance, multiplayer code editor from the creators of Atom
Find the local LLM that actually runs and performs best
LLM inference in C/C++
157 models, 30 providers, one command to find what runs on hardware
Python-free Rust inference server
Supercharge Your LLM with the Fastest KV Cache Layer
Training neural networks on Apple Neural Engine via APIs
A high-performance ML model serving framework, offers dynamic batching
Run AI models locally on your machine with node.js bindings for llama
A GPU-accelerated library containing highly optimized building blocks
C++ image processing and machine learning library with using of SIMD
C++ library for high performance inference on NVIDIA GPUs
LiteRT-LM is Google's production-ready inference framework
lightweight, standalone C++ inference engine for Google's Gemma models
Diffusion model(SD,Flux,Wan,Qwen Image,Z-Image,...) inference
HeavyDB (formerly MapD/OmniSciDB)
3D reconstruction software
Jittor is a high-performance deep learning framework
Fast State-of-the-Art Tokenizers optimized for Research and Production
ArrayFire, a general purpose GPU library
A free, open source, and extensible speech-to-text application