Chao Liu
|
39b829b919
|
adding implicit gemm v3
[ROCm/composable_kernel commit: 1cc683a3a3]
|
2019-05-23 22:10:40 -05:00 |
|
Chao Liu
|
979dc4da2e
|
adding implicit gemm v3
[ROCm/composable_kernel commit: 8a4b59785b]
|
2019-05-22 19:39:56 -05:00 |
|
Chao Liu
|
c562401581
|
behavior has changed (better and worse), figuring out why
[ROCm/composable_kernel commit: 2a48812edb]
|
2019-05-21 16:43:56 -05:00 |
|
Chao Liu
|
45e1ad4dea
|
adding ConstantMergedTensorDescriptor, refactering ConstantTensorDescriptor, Sequence
[ROCm/composable_kernel commit: acd7082fe1]
|
2019-05-21 16:17:58 -05:00 |
|
Chao Liu
|
75df8d78a0
|
rework sequence
[ROCm/composable_kernel commit: a6b95c393b]
|
2019-05-18 23:21:02 -05:00 |
|
Chao Liu
|
4e619d117b
|
rework sequence
[ROCm/composable_kernel commit: df73287b82]
|
2019-05-17 14:56:39 -05:00 |
|
Chao Liu
|
dec8c3ebdd
|
adding implicit gemm v3
[ROCm/composable_kernel commit: 33b5a8556b]
|
2019-05-16 22:23:18 -05:00 |
|
Chao Liu
|
ffd172378a
|
adding implicit gemm v3
[ROCm/composable_kernel commit: 5e5c27a63b]
|
2019-05-16 13:22:40 -05:00 |
|
Chao Liu
|
ac7741cc7c
|
adding implicit gemm v3
[ROCm/composable_kernel commit: b7d052459d]
|
2019-05-15 09:58:17 -05:00 |
|
Chao Liu
|
04e99df5df
|
tuning on vega 20
[ROCm/composable_kernel commit: 2603bb0fe3]
|
2019-04-25 17:28:59 -05:00 |
|
Chao Liu
|
868068eee1
|
implicit gemm v1r3 nchw_cyxk_nkhw
[ROCm/composable_kernel commit: a903146427]
|
2019-04-25 15:14:39 -05:00 |
|
Chao Liu
|
21988c32b4
|
added implicit gemm v1r3 lds_double_buffer NCHW * CYXK = KNHW, reworked static functionals
[ROCm/composable_kernel commit: 569ad66e2a]
|
2019-04-23 17:51:14 -05:00 |
|
Chao Liu
|
d0244d3a51
|
implicit gemm v1r2: adding support for nchw
[ROCm/composable_kernel commit: 19f17df47a]
|
2019-04-18 11:49:09 -05:00 |
|
Chao Liu
|
482e5e9293
|
refactor ConstantTensorDescriptor and functional
[ROCm/composable_kernel commit: 17f3d2d4bc]
|
2019-04-16 17:36:18 -05:00 |
|
Chao Liu
|
8b7eafe959
|
implicit gemm v1r2: only load 1d filter
[ROCm/composable_kernel commit: 00899f191b]
|
2019-04-13 11:19:17 -05:00 |
|
Chao Liu
|
75ca00f748
|
tidy yp
[ROCm/composable_kernel commit: 471830a052]
|
2019-04-09 18:07:36 -05:00 |
|
Chao Liu
|
217b1306e6
|
add more assertion
[ROCm/composable_kernel commit: c075d3f7d9]
|
2019-04-08 12:02:56 -05:00 |
|
Chao Liu
|
d430879858
|
debugging implicit gemm v1: use 10d tensor output
[ROCm/composable_kernel commit: c9fa46af0b]
|
2019-04-08 10:27:32 -05:00 |
|
Chao Liu
|
7cbd63b2d0
|
refactor
[ROCm/composable_kernel commit: e43d7bc63c]
|
2019-04-01 15:17:22 -05:00 |
|
Chao Liu
|
cd883e7581
|
experimenting
[ROCm/composable_kernel commit: 766b0a9eaf]
|
2019-03-24 12:09:57 -05:00 |
|
Chao Liu
|
6fd0910da8
|
refactoring ConstantTensorDescriptor
[ROCm/composable_kernel commit: a0584426ff]
|
2019-03-17 03:22:41 -05:00 |
|
Chao Liu
|
1c962a13ee
|
device_implicit_gemm_convolution_1_chwn_csrk_khwn: use tensor copy (instead of pointwise) for writing output, 3x3 increased from 78% to 84%, 5x5 from 80% to 84%
[ROCm/composable_kernel commit: a65ef90308]
|
2019-02-19 11:47:46 -06:00 |
|
Chao Liu
|
c0baa18a3f
|
change file extension to hip.hpp and hip.cpp
[ROCm/composable_kernel commit: b2888adfbe]
|
2019-02-15 02:13:21 -06:00 |
|