Chao Liu a4f24233e5 manually apply bug fix changes in pr #63 (#64)
* Bug in BlockwiseGemmXdlops_k0mk1_k0nk1_m0n0m1n1m2m3m4n2_v1::MakeCGridDescriptor_M0_N0_M1_N1_M2_M3_M4_N2()
* Bug in ThreadwiseTensorSliceTransfer_v1r3 logic for calculating "forward_sweep"
2021-12-12 18:05:51 -06:00
2021-08-08 17:41:54 +00:00
2021-12-04 16:05:29 -06:00
2021-12-02 20:07:37 -06:00
2021-12-02 20:07:37 -06:00
2021-11-24 12:33:55 -06:00
2018-10-08 22:49:58 -05:00
2021-08-08 17:41:54 +00:00
Description
[DEPRECATED] Moved to ROCm/rocm-libraries repo. NOTE: develop branch is maintained as a read-only mirror
Readme MIT 228 MiB
Languages
C++ 93.1%
Python 4.5%
CMake 1.5%
Shell 0.5%
Pawn 0.2%