Commit Graph

23 Commits

Author SHA1 Message Date
Chao Liu
39b829b919 adding implicit gemm v3
[ROCm/composable_kernel commit: 1cc683a3a3]
2019-05-23 22:10:40 -05:00
Chao Liu
979dc4da2e adding implicit gemm v3
[ROCm/composable_kernel commit: 8a4b59785b]
2019-05-22 19:39:56 -05:00
Chao Liu
c562401581 behavior has changed (better and worse), figuring out why
[ROCm/composable_kernel commit: 2a48812edb]
2019-05-21 16:43:56 -05:00
Chao Liu
45e1ad4dea adding ConstantMergedTensorDescriptor, refactering ConstantTensorDescriptor, Sequence
[ROCm/composable_kernel commit: acd7082fe1]
2019-05-21 16:17:58 -05:00
Chao Liu
75df8d78a0 rework sequence
[ROCm/composable_kernel commit: a6b95c393b]
2019-05-18 23:21:02 -05:00
Chao Liu
4e619d117b rework sequence
[ROCm/composable_kernel commit: df73287b82]
2019-05-17 14:56:39 -05:00
Chao Liu
dec8c3ebdd adding implicit gemm v3
[ROCm/composable_kernel commit: 33b5a8556b]
2019-05-16 22:23:18 -05:00
Chao Liu
ffd172378a adding implicit gemm v3
[ROCm/composable_kernel commit: 5e5c27a63b]
2019-05-16 13:22:40 -05:00
Chao Liu
ac7741cc7c adding implicit gemm v3
[ROCm/composable_kernel commit: b7d052459d]
2019-05-15 09:58:17 -05:00
Chao Liu
04e99df5df tuning on vega 20
[ROCm/composable_kernel commit: 2603bb0fe3]
2019-04-25 17:28:59 -05:00
Chao Liu
868068eee1 implicit gemm v1r3 nchw_cyxk_nkhw
[ROCm/composable_kernel commit: a903146427]
2019-04-25 15:14:39 -05:00
Chao Liu
21988c32b4 added implicit gemm v1r3 lds_double_buffer NCHW * CYXK = KNHW, reworked static functionals
[ROCm/composable_kernel commit: 569ad66e2a]
2019-04-23 17:51:14 -05:00
Chao Liu
d0244d3a51 implicit gemm v1r2: adding support for nchw
[ROCm/composable_kernel commit: 19f17df47a]
2019-04-18 11:49:09 -05:00
Chao Liu
482e5e9293 refactor ConstantTensorDescriptor and functional
[ROCm/composable_kernel commit: 17f3d2d4bc]
2019-04-16 17:36:18 -05:00
Chao Liu
8b7eafe959 implicit gemm v1r2: only load 1d filter
[ROCm/composable_kernel commit: 00899f191b]
2019-04-13 11:19:17 -05:00
Chao Liu
75ca00f748 tidy yp
[ROCm/composable_kernel commit: 471830a052]
2019-04-09 18:07:36 -05:00
Chao Liu
217b1306e6 add more assertion
[ROCm/composable_kernel commit: c075d3f7d9]
2019-04-08 12:02:56 -05:00
Chao Liu
d430879858 debugging implicit gemm v1: use 10d tensor output
[ROCm/composable_kernel commit: c9fa46af0b]
2019-04-08 10:27:32 -05:00
Chao Liu
7cbd63b2d0 refactor
[ROCm/composable_kernel commit: e43d7bc63c]
2019-04-01 15:17:22 -05:00
Chao Liu
cd883e7581 experimenting
[ROCm/composable_kernel commit: 766b0a9eaf]
2019-03-24 12:09:57 -05:00
Chao Liu
6fd0910da8 refactoring ConstantTensorDescriptor
[ROCm/composable_kernel commit: a0584426ff]
2019-03-17 03:22:41 -05:00
Chao Liu
1c962a13ee device_implicit_gemm_convolution_1_chwn_csrk_khwn: use tensor copy (instead of pointwise) for writing output, 3x3 increased from 78% to 84%, 5x5 from 80% to 84%
[ROCm/composable_kernel commit: a65ef90308]
2019-02-19 11:47:46 -06:00
Chao Liu
c0baa18a3f change file extension to hip.hpp and hip.cpp
[ROCm/composable_kernel commit: b2888adfbe]
2019-02-15 02:13:21 -06:00