whisper.cpp/ggml
Masashi Yoshimura 3e0b917514 ggml-webgpu: improve i-quants mul_mat performance and speed up prefill (llama/24530)
* Improve prefill speeds for i-quants

* Fix #if defined() usage in preprocessor guards.
2026-06-19 12:53:43 +03:00
..
cmake ggml : Parallelize quant LUT init (llama/23595) 2026-05-25 12:26:07 +03:00
include Remove padding and multiple D2D copies for MTP (llama/24086) 2026-06-15 10:33:53 +03:00
src ggml-webgpu: improve i-quants mul_mat performance and speed up prefill (llama/24530) 2026-06-19 12:53:43 +03:00
.gitignore
CMakeLists.txt ggml : bump version to 0.15.1 (ggml/1541) 2026-06-15 10:33:53 +03:00