]> git.djapps.eu Git - pkg/ggml/sources/llama.cpp/commit
ggml-cpu: add 128-bit RVV implementation for Quantization Vector Dot (#20633)
authorrehan-10xengineer <redacted>
Thu, 16 Apr 2026 08:15:15 +0000 (13:15 +0500)
committerGitHub <redacted>
Thu, 16 Apr 2026 08:15:15 +0000 (11:15 +0300)
commit1e796eb41fb51950ada45811a303e57a5f4ea974
treef69cb553b02bb63e03131eb846f7aae7247db2e2
parent5637536517ae4ed3eaa22b39c0d479e049097a9b
ggml-cpu: add 128-bit RVV implementation for Quantization Vector Dot (#20633)

* ggml-cpu: add 128-bit impls for i-quants, ternary quants

* ggml-cpu: add 128-bit impls for iq2_xs, iq3_s, iq3_xxs, tq2_0

Co-authored-by: Rehan Qasim <redacted>
* ggml-cpu: refactor; add rvv checks

---------

Co-authored-by: taimur-10x <redacted>
Co-authored-by: Rehan Qasim <redacted>
ggml/src/ggml-cpu/arch/riscv/quants.c