ggml-meta: propagate buffer usage and call init on the new tensors - #27586
Conversation
|
As requested, tested with -sm tensor on a few different model archs on dual RTX A5000. Seems to work great, thanks @max-krasnyansky ! |
|
@youyoulyz Looking good with ggml-metal on my M4 Pro |
|
double-checked on my CUDA setup. @ggerganov should be good to merge |
|
@ggerganov sorry for bugging you. It's blocking for #26501 |
f3aa3ae to
3218261
Compare
No problem at all. I was just planning to add more tests with |
Overview
Another try at fixing
ggml-metato propagate buffer usage and properly initialize new tensors.Original PR was #26502
Tested with
ggml-metalandggml-hexagonbackends.tensor_initmoved aftertensor->datais setup which should resolve issues with CUDA.