llama : enforce the same K and V cache types for DeepSeek V4; enable FA if V cache is quantized - #25871
Merged
ggerganov merged 3 commits intoJul 31, 2026
Merged
background
wait
wait-all
cancel
parallel
Loading