llama.cpp/ggml-cuda.cu at 7930a8a6e89a04c77c51e3ae5dc1cd8e845b6b8f

mirror of https://github.com/ggml-org/llama.cpp.git synced 2025-11-02 09:12:03 +00:00

Files

Johannes Gäßler 3bdc4cd0f5 CUDA: mul_mat_vec_q tiling, refactor mul mat logic (#5434 )

* CUDA: mul_mat_vec_q tiling, refactor mul mat logic

Co-authored-by: slaren <slarengh@gmail.com>

---------

Co-authored-by: slaren <slarengh@gmail.com>

2024-02-11 19:08:39 +01:00

437 KiB

Raw Blame History

View Raw

437 KiB Raw Blame History

437 KiB

Raw Blame History