mb4063 commited on
Commit
7433c9c
·
verified ·
1 Parent(s): 8a04fd1

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +1 -0
README.md CHANGED
@@ -179,6 +179,7 @@ build-rdna4/bin/llama-server \
179
 
180
  To enable MTP speculative decoding (65K context only), add `--spec-type draft-mtp --spec-draft-n-max 3`. See the MTP caveats above before using it for agent work.
181
 
 
182
  ## ROCmFPX bug: prompt-cache checkpoint crash (and fix)
183
 
184
  When `--cache-ram > 0` and `-ctxcp N -cpent N` are used together, the ROCmFPX server may crash on the second request. The bug is in `ggml_backend_tensor_copy` (`ggml/src/ggml-backend.cpp`).
 
179
 
180
  To enable MTP speculative decoding (65K context only), add `--spec-type draft-mtp --spec-draft-n-max 3`. See the MTP caveats above before using it for agent work.
181
 
182
+ # Status (2026-08-15): fixed upstream. Verified on ROCmFPX main @ b2f5829 (build 213): The patch below is only needed for older builds.
183
  ## ROCmFPX bug: prompt-cache checkpoint crash (and fix)
184
 
185
  When `--cache-ram > 0` and `-ctxcp N -cpent N` are used together, the ROCmFPX server may crash on the second request. The bug is in `ggml_backend_tensor_copy` (`ggml/src/ggml-backend.cpp`).