A vector index built on TurboQuant, written in Rust with Python
Implementation of TurboQuant (ICLR 2026)
From-scratch PyTorch implementation of Google's TurboQuant
Accessible large language models via k-bit quantization for PyTorch
Libraries for applying sparsification recipes to neural networks
AIMET is a library that provides advanced quantization and compression
Minimal and clean examples of machine learning algorithms
An implementation of a deep learning recommendation model (DLRM)
A kernel library written in tilelang
Neural Network Compression Framework for enhanced OpenVINO
Open-source large language model family from Tencent Hunyuan
Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs
Machine learning on FPGAs using HLS
A unified library of SOTA model optimization techniques
Pytorch domain library for recommendation systems
Pretrained (Language) Models for Probabilistic Time Series Forecasting
Library to facilitate federated learning research
MiniSom is a minimalistic implementation of the Self Organizing Maps
Z80-μLM is a 2-bit quantized language model
The data structure for multimodal data
Build AI-powered semantic search applications
Build cross-modal and multimodal applications on the cloud
A Python package for extending the official PyTorch
Low-code framework for building custom LLMs, neural networks
Implementation for MatMul-free LM