Radoslav Gerganov
|
58640aa456
|
rpc : better caching of the base buffer pointer (llama/11331)
There is no need to use map, just store the base pointer in the buffer
context.
|
2025-02-03 22:00:57 +02:00 |
Radoslav Gerganov
|
c8c63eeec0
|
rpc : code cleanup (llama/11107)
Remove duplicated macros, use GGML_LOG_ERROR for errors
|
2025-01-14 10:38:01 +02:00 |
matt23654
|
dcbb375779
|
Support for models with non-512-aligned tensors over RPC. (llama/11047)
* Added init tensor calling code
* Added get_alloc_size forwarding
* Cleaned up and improved type/error handling.
* fix: remove trailing whitespaces.
* Cleanup and use GGML error logging functions.
* Handle potentially dangerous edge cases.
* Apply suggestions from code review
Co-authored-by: Diego Devesa <slarengh@gmail.com>
---------
Co-authored-by: Diego Devesa <slarengh@gmail.com>
|
2025-01-14 10:38:01 +02:00 |
Diego Devesa
|
77e3e4a090
|
ggml : add support for dynamic loading of backends (llama/10469)
* ggml : add support for dynamic loading of backends
---------
Co-authored-by: Georgi Gerganov <ggerganov@gmail.com>
|
2024-12-08 20:14:35 +02:00 |
Diego Devesa
|
746bf2596f
|
ggml : build backends as libraries (llama/10256)
* ggml : build backends as libraries
---------
Signed-off-by: Xiaodong Ye <xiaodong.ye@mthreads.com>
Co-authored-by: Georgi Gerganov <ggerganov@gmail.com>
Co-authored-by: R0CKSTAR <xiaodong.ye@mthreads.com>
|
2024-11-20 21:00:08 +02:00 |