whisper.cpp/ggml
Chris Lee 56cb15472d sycl: fix UE4M3 parsing (llama/25608)
The NVFP4 quantization format stores a scaling factor for every group of
16 weights, packed into a single UE4M3 byte.

The SYCL GPU code was converting these scale values using the E4M3 path,
but that's *signed*, and these are unsigned values.
2026-08-07 21:59:49 +03:00
..
cmake ggml : Parallelize quant LUT init (llama/23595) 2026-05-25 12:26:07 +03:00
include mtmd/ggml: add ggml_build_forward_order (llama/26649) 2026-08-07 21:59:49 +03:00
src sycl: fix UE4M3 parsing (llama/25608) 2026-08-07 21:59:49 +03:00
.gitignore whisper : reorganize source code + improve CMake (#2256) 2024-06-26 19:34:09 +03:00
CMakeLists.txt ggml : bump version to 0.18.1 (ggml/1578) 2026-08-04 13:37:47 +03:00