Implementation of TurboQuant (ICLR 2026)
From-scratch PyTorch implementation of Google's TurboQuant
Accessible large language models via k-bit quantization for PyTorch
Libraries for applying sparsification recipes to neural networks
Minimal and clean examples of machine learning algorithms
AIMET is a library that provides advanced quantization and compression
An implementation of a deep learning recommendation model (DLRM)
Neural Network Compression Framework for enhanced OpenVINO
Open-source large language model family from Tencent Hunyuan
Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs
Machine learning on FPGAs using HLS
A unified library of SOTA model optimization techniques
Pytorch domain library for recommendation systems
Pretrained (Language) Models for Probabilistic Time Series Forecasting
Library to facilitate federated learning research
The data structure for multimodal data
Build AI-powered semantic search applications
Z80-μLM is a 2-bit quantized language model
MiniSom is a minimalistic implementation of the Self Organizing Maps
Build cross-modal and multimodal applications on the cloud
Low-code framework for building custom LLMs, neural networks
A Python package for extending the official PyTorch
Big Model Application Development Practice 1
An Open Source implementation of Notebook LM with more flexibility
Implementation for MatMul-free LM