PyTorch/TorchScript/FX compiler for NVIDIA GPUs using TensorRT
Fast Multimodal LLM on Mobile Devices
LiteRT, successor to TensorFlow Lite
Emscripten: An LLVM-to-WebAssembly Compiler
Fast inference engine for Transformer models
Runtime extension of Proximus enabling Deployment on AMD Ryzen™ AI
Deep learning inference framework optimized for mobile platforms