whisper.cpp/ggml
Hongqiang Wang ea8085373a
opencl: add int8 dp4 dense and MoE prefill optimization for Adreno GPUs (llama/25537)
* opencl: add int8 dp4 dense and moe GEMM

* opencl: refactor

---------

Co-authored-by: Li He <lih@qti.qualcomm.com>
2026-07-30 16:31:37 +03:00
..
cmake ggml : Parallelize quant LUT init (llama/23595) 2026-05-25 12:26:07 +03:00
include ggml-et: Initial ET backend (llama/24179) 2026-07-30 16:31:36 +03:00
src opencl: add int8 dp4 dense and MoE prefill optimization for Adreno GPUs (llama/25537) 2026-07-30 16:31:37 +03:00
.gitignore whisper : reorganize source code + improve CMake (#2256) 2024-06-26 19:34:09 +03:00
CMakeLists.txt ggml-et: Initial ET backend (llama/24179) 2026-07-30 16:31:36 +03:00