mirror of
https://github.com/ggml-org/whisper.cpp.git
synced 2026-10-06 22:41:59 +02:00
ggml-cpu: extend RVV quantization vec dot to higher VLENs (llama/22754)
* ggml-cpu: add rvv 512b,1024b impls for iq4_xs * ggml-cpu: refactor; add rvv 512b, 1024b impls for q6_K, i-quants * ggml-cpu: refactor; add 512 and 1024 implementations of tq3_s, iq3_xxs, iq2_s, iq2_xs, iq2_xxs improve iq2_xs impl for rvv 256 Co-authored-by: Rehan Qasim <rehan.qasim@10xengineers.ai> --------- Co-authored-by: taimur-10x <taimur.ahmad@10xengineers.ai> Co-authored-by: Rehan Qasim <rehan.qasim@10xengineers.ai>
This commit is contained in:
committed by
Georgi Gerganov
co-authored by
Rehan Qasim
taimur-10x
parent
00a9728de3
commit
a1a3186887
+3025
-982
File diff suppressed because it is too large
Load Diff
Reference in New Issue
Block a user