tokenspeed-mooncake 0.3.12.post20260725
Source: Pypi.org· July 25, 2026
A KVCache-centric Disaggregated Architecture for large-scale LLM inference and training.
This story was reported by Pypi.org. Read the full original article:
Read on Pypi.org A KVCache-centric Disaggregated Architecture for large-scale LLM inference and training.
Token-Efficient Ingestion Middleware & Concept-Diff Engine for Deep Research AI Agents.
Token-Efficient Ingestion Middleware & Concept-Diff Engine for Deep Research AI Agents.

Token-Efficient Ingestion Middleware & Concept-Diff Engine for Deep Research AI Agents.
Cognitive memory architecture for LLM agents with principled forgetting