llama.cpp

mirror of https://github.com/ggml-org/llama.cpp.git synced 2025-10-27 08:21:30 +00:00

Files

Jeff Bolz 267e99867f vulkan: Use larger loads in scalar/coopmat1 matmul (#15729 )

I think glslang will translate an access like x[i][1].z to
OpAccessChain ... x, i, 1, 2
OpLoad float16_t ...

rather than loading all of x[i] in a single OpLoad. Change the
code to explicitly load the vector/matrix.

2025-09-07 18:53:07 +02:00

cmake

ggml: Skip backend library linking code when GGML_BACKEND_DL=ON (#15094 )

2025-08-07 13:45:41 +02:00

include

ggml-cpu: drop support for nnpa intrinsics (#15821 )

2025-09-06 11:27:28 +08:00

src

vulkan: Use larger loads in scalar/coopmat1 matmul (#15729 )

2025-09-07 18:53:07 +02:00

.gitignore

vulkan : cmake integration (#8119 )

2024-07-13 18:12:39 +02:00

CMakeLists.txt

ggml-cpu: drop support for nnpa intrinsics (#15821 )

2025-09-06 11:27:28 +08:00