]> git.djapps.eu Git - pkg/ggml/sources/whisper.cpp/commit
CUDA: fix thread/block count in quantized cpy kernel launches (llama/26731)
authorRafail Giavrimis <redacted>
Sat, 8 Aug 2026 04:40:04 +0000 (05:40 +0100)
committerGeorgi Gerganov <redacted>
Fri, 14 Aug 2026 19:16:06 +0000 (22:16 +0300)
commit068d3b0b4530506a884fe88aa6e219e870697992
tree6789947da17e0c1b7ebd4fb52a2df3b5065a43aa
parenteb3296f278380600699b9921b136633e6d3a0abe
CUDA: fix thread/block count in quantized cpy kernel launches (llama/26731)

* CUDA: fix thread/block count in quantized cpy kernel launches

* tests: add uneven block count cpy case
ggml/src/ggml-cuda/cpy.cu