ggml webgpu: profiling, CI updates, reworking of command submission (llama/16452)

* Add profiling

* More detailed profiling

* Rework command submission to avoid global locks

* Update wait handling

* try new method of waiting on futures

* Add serializing of command submission in some cases

* Add new pool for timestamp queries and clean up logging

* Serialize command submission in CI and leave a TODO note

* Update webgpu CI

* Add myself as WebGPU codeowner

* Deadlock avoidance

* Leave WebGPU/Vulkan CI serialized

* Fix divide by 0

* Fix logic in division by inflight_threads

* Update CODEOWNERS and remove serialize submit option
This commit is contained in:
Reese Levine
2025-10-12 11:16:23 +03:00
committed by Georgi Gerganov
parent 4bce4fa5e9
commit 4eea3efc49
4 changed files with 491 additions and 242 deletions
File diff suppressed because it is too large Load Diff