whisper.cpp/ggml
Masashi Yoshimura 4b914096e8 ggml-webgpu: tune subgroup split (d_split) in flash_attn_vec (llama/25418) 2026-07-10 13:06:42 +03:00
..
cmake ggml : Parallelize quant LUT init (llama/23595) 2026-05-25 12:26:07 +03:00
include Add Q2_0 quantization: type definition and CPU backend (llama/24448) 2026-07-10 13:06:42 +03:00
src ggml-webgpu: tune subgroup split (d_split) in flash_attn_vec (llama/25418) 2026-07-10 13:06:42 +03:00
.gitignore
CMakeLists.txt ggml : bump version to 0.15.3 (ggml/1550) 2026-06-26 16:03:57 +03:00