YaelGitAccount
|
851553ea6b
|
cuda: add SET operation support (#16804)
* feat(cuda): add GGML_OP_SET support
Implement CUDA kernel for SET operation with f32 support.
All tests passing (14598/14598).
* cuda(set): add I32 support; keep F32
* refactor(cuda): use ggml_cuda_cpy to unify SET operator logic and remove code duplication
* Update ggml/src/ggml-cuda/ggml-cuda.cu
Co-authored-by: Sigbjørn Skjæret <sigbjorn.skjaeret@scala.com>
* Update ggml/src/ggml-cuda/set.cu
Co-authored-by: Sigbjørn Skjæret <sigbjorn.skjaeret@scala.com>
---------
Co-authored-by: Sigbjørn Skjæret <sigbjorn.skjaeret@scala.com>
|
2025-10-28 20:10:28 +01:00 |
|