Describe SparkCache behavior and qualification directly - #34
Closed
FujitsuPolycom wants to merge 1 commit into
Closed
Describe SparkCache behavior and qualification directly#34FujitsuPolycom wants to merge 1 commit into
FujitsuPolycom wants to merge 1 commit into
Conversation
Replace the lifecycle warning label with an exact qualification boundary and describe subsequent-request reuse without temporal shorthand. Documentation only; cache identity and runtime behavior are unchanged. Validation: Markdown diff checks passed.
Owner
Author
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
Resulting documentation
The README begins with SparkCache as a persistent, rank-local NVMe context cache that reuses the longest verified stored prefix. State that cannot prove identity, compatibility, all-rank availability, and payload integrity is rejected, and vLLM computes the request normally.
The capability table distinguishes:
Artifact details are grouped under a dedicated qualification section instead of occupying the opening. The performance section highlights restore size and concurrency effects without reproducing benchmark methodology. Operations describe the default full-snapshot schema, opt-in copy-on-write tails, SSD-endurance limitations, and unsupported write-budget controls.
User-facing prose no longer calls the accelerated path native restore or native placement. Historical filenames and the
sparkcache/native/source directory remain unchanged because they are durable repository identifiers.Compatibility
Documentation only. Cache identity values, digest salts, chunk geometry, stored schemas, runtime behavior, and the SparkCache CUDA placement ABI are unchanged. Cache namespace impact: none.
Validation
This draft is stacked on the qualified SparkCache behavior consolidation in PR #32 and supersedes the earlier README draft in PR #24.