ik_llama.cpp/tests/test-backend-ops.cpp at 584ab7fbfba1954703fba7c6c1636c7f20cf35e9

mirror of https://github.com/ikawrakow/ik_llama.cpp.git synced 2026-03-02 18:10:02 +00:00

Files

Johannes Gäßler 871641d19e CUDA: add FP32 FlashAttention vector kernel (#7188 )

* CUDA: add FP32 FlashAttention vector kernel

* fixup! CUDA: add FP32 FlashAttention vector kernel

* fixup! fixup! CUDA: add FP32 FlashAttention vector kernel

* fixup! fixup! fixup! CUDA: add FP32 FlashAttention vector kernel

2024-05-12 19:40:45 +02:00

80 KiB

Raw Blame History

View Raw

80 KiB Raw Blame History

80 KiB

Raw Blame History