ggml-cpu: add 128-bit RVV implementation for Quantization Vector Dot (llama/20633)

* ggml-cpu: add 128-bit impls for i-quants, ternary quants

* ggml-cpu: add 128-bit impls for iq2_xs, iq3_s, iq3_xxs, tq2_0

Co-authored-by: Rehan Qasim <rehan.qasim@10xengineers.ai>

* ggml-cpu: refactor; add rvv checks

---------

Co-authored-by: taimur-10x <taimur.ahmad@10xengineers.ai>
Co-authored-by: Rehan Qasim <rehan.qasim@10xengineers.ai>
This commit is contained in:
rehan-10xengineer
2026-04-30 11:29:11 +03:00
committed by Georgi Gerganov
co-authored by Rehan Qasim taimur-10x
parent 07c181b57f
commit 94d6d0b743
File diff suppressed because it is too large Load Diff