Related Products
|
||||||
About
LMCache is an open source Knowledge Delivery Network (KDN) designed as a caching layer for large language model serving that accelerates inference by reusing KV (key-value) caches across repeated or overlapping computations. It enables fast prompt caching, allowing LLMs to “prefill” recurring text only once and then reuse those stored KV caches, even in non-prefix positions, across multiple serving instances. This approach reduces time to first token, saves GPU cycles, and increases throughput in scenarios such as multi-round question answering or retrieval augmented generation. LMCache supports KV cache offloading (moving cache from GPU to CPU or disk), cache sharing across instances, and disaggregated prefill, which separates the prefill and decoding phases for resource efficiency. It is compatible with inference engines like vLLM and TGI and supports compressed storage, blending techniques to merge caches, and multiple backend storage options.
|
About
Terracotta Server Platform is a distributed in-memory data management system that supports Terracotta products such as Ehcache and TCStore. The platform acts as the backbone for Terracotta clusters and helps add distributed caching capabilities to Ehcache deployments. A Terracotta Server Array can range from a basic two-node setup to a larger multi-node configuration for greater scale, performance, and failover coverage. Terracotta Server provides features such as distributed in-memory data management, simple scalability, high availability, health monitoring, and automatic node reconnection. It can manage significantly more in-memory data than traditional data grids while allowing teams to expand server instances as demand grows. Terracotta Server Platform is designed for development and infrastructure teams that need reliable distributed caching, clustered data management, and high-performance in-memory storage.
|
|||||
Platforms Supported
Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook
|
Platforms Supported
Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook
|
|||||
Audience
AI engineers and infrastructure teams looking for a tool to lower latency, reduce compute cost, and scale throughput
|
Audience
Terracotta Server Platform is best suited for Java developers, infrastructure teams, enterprise application teams, and organizations that need distributed caching, clustered Ehcache support, in-memory data management, high availability, and scalable server-side cache storage
|
|||||
Support
Phone Support
24/7 Live Support
Online
|
Support
Phone Support
24/7 Live Support
Online
|
|||||
API
Offers API
|
API
Offers API
|
|||||
Screenshots and Videos |
Screenshots and Videos |
|||||
Pricing
Free
Free Version
Free Trial
|
Pricing
No information available.
Free Version
Free Trial
|
|||||
Reviews/
|
Reviews/
|
|||||
Training
Documentation
Webinars
Live Online
In Person
|
Training
Documentation
Webinars
Live Online
In Person
|
|||||
Company InformationLMCache
United States
lmcache.ai/
|
Company InformationTerracotta
www.terracotta.org
|
|||||
Alternatives |
Alternatives |
|||||
|
|
||||||
|
|
||||||
|
|
|
|||||
Categories |
Categories |
|||||
Integrations
FF4J
|
||||||
|
|
|