A fast cache that automatically deletes the least recently used items
Nothing but Cache
Free RAM cleaner uses native Windows features to optimize memory areas
Trainable latent-memory framework for 100M-token contexts
Go cache library that brings you multiple ways to manage caches
A high performance memory-bound Go cache
JetCache is a Java cache framework
In-memory and distributed caching toolkit for Elixir
Redundancy-aware KV Cache Compression for Reasoning Models
DeepSeek-native AI coding agent for your terminal
A 2.78-trillion-parameter Kimi K3 running inference on a single CPU
Unified KV Cache Compression Methods for Auto-Regressive Models
FusionCache is an easy to use, fast and robust cache
Open-source TypeScript terminal coding agent for DeepSeek-V4
Image loading system
Gemma 4 26B-A4B inference in ~2 GB of RAM on any M-series MacBook
A high performance caching library for Java
Distributed, in-memory key/value store and cache
User space software for Intel(R) Resource Director Technology
An expressive, efficient attention architecture
Mooncake is the serving platform for Kimi
FlashMLA: Efficient Multi-head Latent Attention Kernels
Bentocache is a robust multi-tier caching library for Node.js app
Lightweight, pure-Swift library for downloading images from the web
Zen Patched Kernel Sources