OpenCL Token Generation Acceleration (#1459)

* Move back to C++ for OpenCL * Refactor OpenCL code to work more like the CUDA code, add missing functions * Deduplicate dequant kernels * Add OpenCL compile options * Use compile args for preprocessing constants * Restore default platform + device selection by id behavior --------- Co-authored-by: Johannes Gäßler <johannesg@5d6.de> Co-authored-by: Henri Vasserman <henv@hot.ee>
2025-06-26 19:55:04 +00:00 · 2023-05-22 23:33:24 +02:00
parent 7e4ea5beff
commit 2e6cd4b025
8 changed files with 1113 additions and 536 deletions
--- a/CMakeLists.txt
+++ b/CMakeLists.txt
@ -201,7 +201,7 @@ if (LLAMA_CLBLAST)
    if (CLBlast_FOUND)
        message(STATUS "CLBlast found")

-        set(GGML_OPENCL_SOURCES ggml-opencl.c ggml-opencl.h)
+        set(GGML_OPENCL_SOURCES ggml-opencl.cpp ggml-opencl.h)

        add_compile_definitions(GGML_USE_CLBLAST)