Files
composable_kernel/include/ck_tile/core
Yung-sheng Tu 604c56bc0e [rocm-libraries] ROCm/rocm-libraries#7850 (commit e8f2756)
=?UTF-8?q?style:=20[CK=20TILE]=20Unification=20Work=20?=
 =?UTF-8?q?=E2=80=93=20Unify=20format=20MFMA=20part=20(#7850)?=
MIME-Version: 1.0
Content-Type: text/plain; charset=UTF-8
Content-Transfer-Encoding: 8bit

## Motivation

This PR unifies the parameter comments and simplifies the docs for
`amdgcn_mma` specialisations of `MfmaOp`.

## Technical Details

Except for the two things mentioned above, it also simplifies the sparse
traits, unifies the usages of `enable_if_target_id_t`, and cleans up the
files in
[include/ck_tile/core/arch/mma](https://github.com/ROCm/rocm-libraries/tree/users/yungshengtu/ck/unification/unify_format_mfma/projects/composablekernel/include/ck_tile/core/arch/mma).

**NOTE: The first commit is not in the scope of this PR.**

## Test Plan

Test has existed.

## Test Result

Test should pass.

## Submission Checklist

- [x] Look over the contributing guidelines at
https://github.com/ROCm/ROCm/blob/develop/CONTRIBUTING.md#pull-requests.

close #8907
2026-06-29 18:51:17 +00:00
..
2024-04-15 19:27:12 -05:00

ck_tile/core

ck_tile/core contains every basic functions and structures to create a GPU kernel using ck_tile. User should only include ck_tile/core.hpp this single header to use all the functionality. Everything is under ck_tile namespace. The coding style under this folder should be similar to std (snake_case for structure/function, Camel for template types...)

algorithm/
    coordinate transform and some other reusable algorithm
arch/
    contains some basic device building block like mma, buffer addressing, etc...
container/
    contains basic container data structure, array/sequence/tuple/...
numeric/
    data type, and data type related math
tensor/
    tensor descriptors and tile level API
utility/
    other utility function for both host/device