chore: ⬆️ Update TheTom/llama-cpp-turboquant to c26cbdffcf6fc9b7430cd6b117757e9a3f70b7ea#10235
chore: ⬆️ Update TheTom/llama-cpp-turboquant to c26cbdffcf6fc9b7430cd6b117757e9a3f70b7ea#10235localai-bot wants to merge 1 commit into
c26cbdffcf6fc9b7430cd6b117757e9a3f70b7ea#10235Conversation
33783df to
95d0d4f
Compare
73eb521daebc85da7c91d37178940b99a5524cf6f0d5cc9f71babf7cba8679a06efd602d26b70b06
7c48dfd to
2467c24
Compare
f0d5cc9f71babf7cba8679a06efd602d26b70b067985f6b90bf19881ab7c7a8444954e91cae36056
ba90636 to
b2603de
Compare
7985f6b90bf19881ab7c7a8444954e91cae3605635ac80d55b8e8e2a0d179413caddc2622b54c248
b2603de to
b863245
Compare
35ac80d55b8e8e2a0d179413caddc2622b54c2484595fff0bbd15ee01663699b788eea70e7e1cd69
fe42778 to
beacee4
Compare
4595fff0bbd15ee01663699b788eea70e7e1cd69a33ef00b13476e9c609caecc3c1c015b8615011d
a21dc73 to
2de5c5b
Compare
…10235) The fork a33ef00b HIP-ported ggml_cuda_copy2d_across_devices() itself (it now guards the cudaMemcpy3DPeer fast path with #if !defined(GGML_USE_HIP) && !defined(GGML_USE_MUSA) and falls back to a plain 2D device-to-device copy), so the former copy2d hunk's anchors went stale. Retire that hunk and re-anchor the still-needed event hunk: ggml_backend_cuda_device_event_new() still uses plain cudaEventCreate, which ggml's HIP shim does not alias, so keep the cudaEventCreateWithFlags (cudaEventDisableTiming) rewrite. Validated with apply-patches.sh against a fresh clone at the pinned commit. Signed-off-by: Ettore Di Giacinto <mudler@localai.io> Assisted-by: Claude:claude-opus-4-8 [Claude Code]
fd438a1 to
787c635
Compare
…10235) The fork a33ef00b HIP-ported ggml_cuda_copy2d_across_devices() itself (it now guards the cudaMemcpy3DPeer fast path with #if !defined(GGML_USE_HIP) && !defined(GGML_USE_MUSA) and falls back to a plain 2D device-to-device copy), so the former copy2d hunk's anchors went stale. Retire that hunk and re-anchor the still-needed event hunk: ggml_backend_cuda_device_event_new() still uses plain cudaEventCreate, which ggml's HIP shim does not alias, so keep the cudaEventCreateWithFlags (cudaEventDisableTiming) rewrite. Validated with apply-patches.sh against a fresh clone at the pinned commit. Signed-off-by: Ettore Di Giacinto <mudler@localai.io> Assisted-by: Claude:claude-opus-4-8 [Claude Code]
db5dabc to
b730ea9
Compare
a33ef00b13476e9c609caecc3c1c015b8615011d0c8fcfe737ac7cfa79a5b8558024fd7ede79010f
b730ea9 to
7a1fcf7
Compare
2e78b9b to
88a3c62
Compare
558c6b78e4f8cf92ec19539ff89b6d13f4183feb3d42fa1eb6e9f35c80ac86084637f574c3c764db
9777cc8 to
08694ce
Compare
3d42fa1eb6e9f35c80ac86084637f574c3c764db206e18efc2966e8f0176e6f7012f8f9610195937
08694ce to
6f316dd
Compare
206e18efc2966e8f0176e6f7012f8f9610195937c3e6dbb13d40e2e42f7a964bd5d745fbf86e4495
f4319bb to
8d7cfce
Compare
c3e6dbb13d40e2e42f7a964bd5d745fbf86e44954503343ffc05c09f6b50c309c8ecbabb49c66ea2
92f2c17 to
2fd33b5
Compare
4503343ffc05c09f6b50c309c8ecbabb49c66ea2f27268914e948aec208937e8b867a68d85f886ce
fa87b16 to
edf201b
Compare
|
The current red matrix is a real patch-refresh failure, not infra: upstream commit |
edf201b to
568cf0f
Compare
f27268914e948aec208937e8b867a68d85f886ce471fb4ec8fc75bbebbcb81d4d6b5e260e9c6e831
568cf0f to
efc2404
Compare
471fb4ec8fc75bbebbcb81d4d6b5e260e9c6e831c26cbdffcf6fc9b7430cd6b117757e9a3f70b7ea
965bd90 to
4b35fc2
Compare
Signed-off-by: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>
4b35fc2 to
7d6bfb4
Compare
|
@localai-org-maint-bot supersede this one in a new pr |
Changes: https://github.com/TheTom/llama-cpp-turboquant/compare/7d9715f1f071fa07c7b2ad3dbfd320b314139e65..c26cbdffcf6fc9b7430cd6b117757e9a3f70b7ea