Piotr Wilkin (ilintar) and Johannes Gäßler
4d506f58e4
CUDA: replace GGML_FA_ALL_QUANTS with GGML_FA_QUANTS, more control over what is compiled (llama/28079)
...
* CUDA: add configurable FA quant combinations
Assisted-by: Codex
* remove all flags but , add runtime fallback with warning for uncompiled combination
* Update docs/build.md
Co-authored-by: Johannes Gäßler <johannesg@5d6.de >
* apply code review comments
---------
Co-authored-by: Johannes Gäßler <johannesg@5d6.de >
2026-09-14 20:45:06 +03:00
xctan
0c25129d30
ggml-cpu : rework weak alias on apple targets (llama/14146)
...
* ggml-cpu : rework weak alias on apple targets
* fix powerpc detection
* fix ppc detection
* fix powerpc detection on darwin
2025-06-18 12:40:34 +03:00
Christian Kastner
1d7b3c79f4
cmake: Factor out CPU architecture detection (llama/13883)
...
* cmake: Define function for querying architecture
The tests and results match exactly those of src/CMakeLists.txt
* Switch arch detection over to new function
2025-06-01 15:14:44 +03:00
Georgi Gerganov
8ca67df291
ggml : sync/merge cmake,riscv,powerpc, add common.cmake (ggml/0)
2025-03-27 11:06:03 +02:00