llama.cpp/ggml-cuda/common.cuh at 3292733f95d4632a956890a438af5192e7031c12

mirror of https://github.com/ggml-org/llama.cpp.git synced 2025-11-01 09:01:57 +00:00

Files

Johannes Gäßler a743d76a01 CUDA: generalize FP16 fattn vec kernel (#7061 )

* CUDA: generalize FP16 fattn vec kernel

* disable unsupported head sizes for AMD in test

* try AMD fix

* fix batch size 2-8

* partially revert changes

2024-05-09 14:32:02 +02:00

23 KiB

Raw Blame History

View Raw

23 KiB Raw Blame History

23 KiB

Raw Blame History