Vulkan Optimizations and Fixes (#8959)

mirror of https://github.com/ggml-org/llama.cpp.git synced 2025-06-28 12:25:03 +00:00

* Optimize Vulkan REPEAT performance

* Use Vulkan GLSL fused multiply-add instruction where possible

* Add GGML_VULKAN_PERF option to output performance data per operator

* Rework and fix Vulkan descriptor set and descriptor pool handling

* Fix float32 concat f16 shader validation error

* Add Vulkan GROUP_NORM eps parameter

* Fix validation error with transfer queue memory barrier flags

* Remove trailing whitespaces

This commit is contained in:

0cc4m

2024-08-14 18:32:53 +02:00

committed by

GitHub

parent 98a532d474

commit 5fd89a70ea

16 changed files with 781 additions and 851 deletions

1388

ggml/src/ggml-vulkan.cpp

View File

File diff suppressed because it is too large Load Diff

Vulkan Optimizations and Fixes (#8959)

1388 ggml/src/ggml-vulkan.cpp View File

1388

ggml/src/ggml-vulkan.cpp

View File