Files
llama.cpp/ggml-cuda/template-instances/mmq-instance-q2_k.cu
Johannes Gäßler 7d1a378b8f CUDA: refactor mmq, dmmv, mmvq (#7716)
* CUDA: refactor mmq, dmmv, mmvq

* fix out-of-bounds write

* struct for qk, qr, qi

* fix cmake build

* mmq_type_traits
2024-06-05 16:53:00 +02:00

138 B