User-friendly AI Interface
Port of OpenAI's Whisper model in C/C++
Port of Facebook's LLaMA model in C/C++
Run Local LLMs on Any Device. Open-source
A high-throughput and memory-efficient inference and serving engine
ONNX Runtime: cross-platform, high performance ML inferencing
OpenVINO™ Toolkit repository
The free, Open Source alternative to OpenAI, Claude and others
High-performance neural network inference framework for mobile
Open-Source AI Camera. Empower any camera/CCTV
Protect and discover secrets using Gitleaks
Pure C++ implementation of several models for real-time chatting
Operating LLMs in production
Neural Network Compression Framework for enhanced OpenVINO
Official inference library for Mistral models
Efficient few-shot learning with Sentence Transformers
LLMs as Copilots for Theorem Proving in Lean
An MLOps framework to package, deploy, monitor and manage models
LLM training code for MosaicML foundation models
On-device AI across mobile, embedded and edge for PyTorch
Everything you need to build state-of-the-art foundation models
LLM.swift is a simple and readable library
Optimizing inference proxy for LLMs
Build Production-ready Agentic Workflow with Natural Language
The official Python client for the Huggingface Hub