Skip to content

ggml : fix ggml_backend_buft_get_alloc_size() guard - #28038

Merged
ggerganov merged 1 commit into
masterfrom
gg/ggml-may-expand-fix
Aug 30, 2026
Merged

ggml : fix ggml_backend_buft_get_alloc_size() guard#28038
ggerganov merged 1 commit into
masterfrom
gg/ggml-may-expand-fix

Conversation

@ggerganov

Copy link
Copy Markdown
Member

Overview

cont #27960

Didn't take into account that CUDA pads quantized tensors.

Fix: https://github.com/ggml-org/llama.cpp/actions/runs/33296587275/job/99217196889#step:3:31473

Requirements

@ggerganov
ggerganov requested a review from a team as a code owner August 30, 2026 16:42
@github-actions github-actions Bot added ggml changes relating to the ggml tensor library for machine learning CUDA Related to the CUDA backend labels Aug 30, 2026
@ggerganov
ggerganov merged commit 6d1479c into master Aug 30, 2026
26 of 30 checks passed
@ggerganov
ggerganov deleted the gg/ggml-may-expand-fix branch August 30, 2026 17:25
fewtarius pushed a commit to fewtarius/CachyLLama that referenced this pull request Sep 5, 2026
thecodacus pushed a commit to thecodacus/llama.cpp that referenced this pull request Sep 7, 2026
SteelPh0enix pushed a commit to SteelPh0enix/llama.cpp-qwen4exp that referenced this pull request Sep 8, 2026
SteelPh0enix pushed a commit to SteelPh0enix/llama.cpp-qwen4exp that referenced this pull request Sep 8, 2026
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

CUDA Related to the CUDA backend ggml changes relating to the ggml tensor library for machine learning

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant