For the complete documentation index, see llms.txt. Markdown versions of all pages are available by appending .md to any URL (e.g. /get-started.md).
Python module
max.pipelines.kv_cache.paged_kv_cache.kv_group_coordinator
Cache group coordinators: one per set of leaves written in lockstep.
Cache group coordinators
FullKVGroupCoordinator | A group whose caches read their whole history. |
|---|---|
KVGroupCoordinatorInterface | Finds and claims the prefix-cache hit one group of caches can serve. |
SlidingWindowKVGroupCoordinator | A group that needs blocks_in_window consecutive blocks for a hit. |