whisper.cpp/ggml
Ruben Ortlam ba36235565 vulkan: use flops instead of weight tensor size for submission heuristic (llama/25005)
* vulkan: extract flops calculation into function

* use flops instead of matmul src0 tensor size for submission threshold

* use unsigned ints
2026-07-10 13:06:42 +03:00
..
cmake ggml : Parallelize quant LUT init (llama/23595) 2026-05-25 12:26:07 +03:00
include sycl : support --split-mode tensor (llama/24152) 2026-06-26 16:03:57 +03:00
src vulkan: use flops instead of weight tensor size for submission heuristic (llama/25005) 2026-07-10 13:06:42 +03:00
.gitignore whisper : reorganize source code + improve CMake (#2256) 2024-06-26 19:34:09 +03:00
CMakeLists.txt ggml : bump version to 0.15.3 (ggml/1550) 2026-06-26 16:03:57 +03:00