whisper.cpp/ggml/src
Georgi Gerganov d201705e71 metal : fix mul-mm condition + fix mul-mv permuted kernels (llama/16494) 2025-10-12 11:16:23 +03:00
..
ggml-blas rename optimize_graph to graph_optimize (llama/16082) 2025-09-20 13:46:39 +03:00
ggml-cann CANN: Improve ACL graph matching (llama/16166) 2025-10-12 11:16:23 +03:00
ggml-cpu cpu : optimize the ggml NORM operation (llama/15953) 2025-10-12 11:16:23 +03:00
ggml-cuda cuda : avoid initializing unused devices (llama/16510) 2025-10-12 11:16:23 +03:00
ggml-hip HIP: Disable ROCWMMA fattn on CDNA when compiled against ROCWMMA 2.0.0 (llama/16221) 2025-10-12 11:16:23 +03:00
ggml-metal metal : fix mul-mm condition + fix mul-mv permuted kernels (llama/16494) 2025-10-12 11:16:23 +03:00
ggml-musa musa: update compile flags (llama/16265) 2025-10-12 11:16:23 +03:00
ggml-opencl opencl: support pad_ext (llama/15888) 2025-10-12 11:16:23 +03:00
ggml-rpc rpc : check src buffer when copying tensor (llama/16421) 2025-10-12 11:16:23 +03:00
ggml-sycl refactor soft_max, add soft_max_back (llama/16472) 2025-10-12 11:16:23 +03:00
ggml-vulkan vulkan: use a more appropriate amount of threads when generating shaders (llama/16418) 2025-10-12 11:16:23 +03:00
ggml-webgpu ggml webgpu: profiling, CI updates, reworking of command submission (llama/16452) 2025-10-12 11:16:23 +03:00
ggml-zdnn zdnn: refactor codebase + add docs (llama/16178) 2025-09-29 15:18:09 +03:00
CMakeLists.txt cmake : Dont define XOPENSOURCE on AIX (llama/16481) 2025-10-12 11:16:23 +03:00
ggml-alloc.c ggml : fix graph reallocation with multiple chunks (llama/16396) 2025-10-12 11:16:23 +03:00
ggml-backend-impl.h rpc : add support for multiple devices (llama/16276) 2025-10-12 11:16:23 +03:00
ggml-backend-reg.cpp ggml-backend : add root cause in error message if loading backend library fails (llama/16172) 2025-09-30 12:31:00 +03:00
ggml-backend.cpp llama: print memory breakdown on exit (llama/15860) 2025-09-29 15:18:10 +03:00
ggml-common.h llama : add gpt-oss (llama/15091) 2025-08-18 20:30:45 +03:00
ggml-impl.h model : Apertus model implementation (llama/15852) 2025-10-12 11:16:23 +03:00
ggml-opt.cpp finetune: SGD optimizer, more CLI args (llama/13873) 2025-08-18 20:30:45 +03:00
ggml-quants.c ggml : fix uninitialized is_on_grid in quantize_row_iq3_xxs_impl (llama/15928) 2025-09-29 15:18:09 +03:00
ggml-quants.h llama : add gpt-oss (llama/15091) 2025-08-18 20:30:45 +03:00
ggml-threading.cpp ggml : build backends as libraries (llama/10256) 2024-11-20 21:00:08 +02:00
ggml-threading.h remove CMAKE_WINDOWS_EXPORT_ALL_SYMBOLS (llama/10797) 2024-12-18 12:52:16 +02:00
ggml.c ggml webgpu: add support for soft_max, optimize rms_norm (llama/16357) 2025-10-12 11:16:23 +03:00
ggml.cpp ggml : Print backtrace on uncaught C++ exceptions (ggml/1232) 2025-05-29 09:56:26 +03:00
gguf.cpp ggml : prevent integer overflow in gguf tensor size calculation (llama/14595) 2025-07-12 19:23:56 +03:00