Skip to content

Bind DFlash7 to canonical SparkCache CUDA keys - #137

Closed
FujitsuPolycom wants to merge 1 commit into
codex/glm53-dflash7-draft-loader-integrationfrom
codex/glm53-dflash7-canonical-cuda-keys
Closed

Bind DFlash7 to canonical SparkCache CUDA keys#137
FujitsuPolycom wants to merge 1 commit into
codex/glm53-dflash7-draft-loader-integrationfrom
codex/glm53-dflash7-canonical-cuda-keys

Conversation

@FujitsuPolycom

Copy link
Copy Markdown
Owner

Result

Advances the exact DFlash7 image and both loader profiles from SparkCache PR #25 to PR #26 canonical CUDA configuration.

The image pins SparkCache commit b95aa8ab0068dc66a6892a5c311d7e9dd4a9c55a, Git tree 723fc604d73911a9e907798fd7932e4fc9c95df5, and independently reproduced clean-source SHA-256 48e008ba0cbd12f1ffae1c28388ea83310f41c6219c955e13d63ab171290d8de. Prepared image labels and the verifier require org.sparkcache.cuda-config-schema=canonical-v1.

Configuration

Both DFlash7 profiles use only the canonical SparkCache CUDA restore, placement-library, digest, arena, and restore-worker keys. The current launcher can consume either profile directly; no PR25 compatibility profile or legacy-key rewrite is required.

The DFlash draft-loader patch, vLLM native/Python identities, B12X, checkpoint identities, DFlash depth, draft TP4, page-tail behavior, and vLLM ownership contract remain unchanged.

Cache compatibility

No cache namespace change. Checkpoint identities, publication schema, record vocabulary, digest salts, TP/DCP geometry, vLLM patch bytes, lease contract, and CUDA placement ABI are unchanged. Compatible PR25 page-tail entries remain eligible.

Validation

  • 29 focused GPU-free tests passed.
  • Exact prepared context reproduced PR26 commit, tree, source digest, runtime patch receipt, and canonical image labels.
  • Maintained suite: 1,964 passed, 9 skipped, one unrelated stacked-base README assertion failed.
  • Ruff, Python compilation, JSON parsing, Bash syntax, diff, and prose checks passed.

No image was built or published, and no service was modified.

Advance the exact DFlash7 image and both loader profiles to SparkCache commit b95aa8ab, Git tree 723fc604, and independently reproduced deployable-source digest 48e008ba. Prepared image labels and verification now require the canonical-v1 CUDA configuration schema.

Profiles use only spark_cache_cuda_restore, CUDA placement library, digest, arena, and restore-worker keys. No PR25 compatibility profile or legacy-key rewrite is required by the launcher path.

Cache compatibility: checkpoint identities, page-tail publication schema, record vocabulary, digest salts, TP/DCP geometry, vLLM patch bytes, lease contract, and CUDA placement ABI are unchanged. Compatible PR25 page-tail entries remain in the same namespace.

Validation: 29 focused tests passed; exact context preparation reproduced the PR26 commit, tree, source digest, runtime patch receipt, and canonical image labels; maintained suite reported 1,964 passed and 9 skipped with one unrelated stacked-base README assertion failure; Ruff, Python, JSON, Bash, diff, and prose checks passed.
@FujitsuPolycom

Copy link
Copy Markdown
Owner Author

The resulting runtime, operator contract, and evidence are consolidated in retained draft stack #146#147#150. Independent page-base research remains in #149. Closing this superseded draft and deleting only its remote head branch.

@FujitsuPolycom
FujitsuPolycom deleted the codex/glm53-dflash7-canonical-cuda-keys branch August 31, 2026 01:55
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant