ggml : fix quant dot product with odd number of blocks (#8549)

mirror of https://github.com/ggml-org/llama.cpp.git synced 2025-06-28 20:25:20 +00:00

* ggml : fix iq4_nl dot product with odd number of blocks

* ggml : fix odd blocks for ARM_NEON (#8556)

* ggml : fix iq4_nl dot product with odd number of blocks

* ggml : fix q4_1

* ggml : fix q5_0

* ggml : fix q5_1

* ggml : fix iq4_nl metal

ggml-ci

* ggml : fix q4_0

* ggml : fix q8_0

ggml-ci

* ggml : remove special Q4_0 code for first 2 blocks

* ggml : fix sumf redefinition

---------

Co-authored-by: slaren <slarengh@gmail.com>

---------

Co-authored-by: Georgi Gerganov <ggerganov@gmail.com>

This commit is contained in:

slaren

2024-07-19 17:17:27 +02:00

committed by

GitHub

parent 57b1d4f9eb

commit 87e397d00b

4 changed files with 364 additions and 503 deletions

832

ggml/src/ggml-quants.c

View File

File diff suppressed because it is too large Load Diff

ggml : fix quant dot product with odd number of blocks (#8549)

832 ggml/src/ggml-quants.c View File

832

ggml/src/ggml-quants.c

View File