| 文件 | 最后提交记录 | 最后更新时间 |
|---|---|---|
[mlir][AMDGPU] Enable emulating vector buffer_atomic_fadd for bf16 on gfx942 (#129029) - Change to make sure architectures < gfx950 emulate bf16 buffer_atomic_fadd - Add tests for bf16 buffer_atomic_fadd and architectures: gfx12, gfx942 and gfx950 --------- Co-authored-by: Jakub Kuderski <kubakuderski@gmail.com> | 1 年前 | |
[mlir][amdgpu] fold memref.subview/expand_shape/collapse_shape into amdgpu.gather_to_lds for DST operand (#152277) | 11 个月前 | |
[mlir][AMDGPU] Plumb address space 7 through MLIR, add address_space attr. (#125594) This commit adds support for casting memrefs into fat raw buffer pointers to the AMDGPU dialect. Fat raw buffer pointers - or, in LLVM terms, ptr addrspcae(7), allow encapsulating a buffer descriptor (as produced by the make.buffer.rsrc intrinsic or provided from some API) into a pointer that supports ordinary pointer operations like load or store. This allows people to take advantage of the additional semantics that buffer_load and similar instructions provide without forcing the use of entirely separate amdgpu.raw_buffer_* operations. Operations on fat raw buffer pointers are translated to the corresponding LLVM intrinsics by the backend. This commit also goes and and defines a #amdgpu.address_space<> attribute so that AMDGPU-specific memory spaces can be represented. Only #amdgpu.address_space<fat_raw_buffer> will work correctly with the memref dialect, but the other possible address spaces are included for completeness. --------- Co-authored-by: Jakub Kuderski <kubakuderski@gmail.com> Co-authored-by: Prashant Kumar <pk5561@gmail.com> | 1 年前 | |
[mlir][amdgpu] Update scaled_mfma assembly format with intrinsic shape (#165044) Use the same format as introduced for wmma by https://github.com/llvm/llvm-project/pull/164920 and for mfma by https://github.com/llvm/llvm-project/pull/165037. | 9 个月前 | |
[mlir][amdgpu] Add Inliner interface (#162873) All the amdgpu dialect ops can be inlined. --------- Signed-off-by: Ivan Butygin <ivan.butygin@gmail.com> | 9 个月前 | |
[mlir][amdgpu] Add lowerings for ScaledExtPacked816 (#168123) * Adds lowerings for amdgpy.scaled_ext_packed816 * updates verifiers | 8 个月前 | |
[mlir][AMDGPU] Add better load/store lowering for full mask (#146748) This patch adds a better maskedload/maskedstore lowering on amdgpu backend for loads which are either fully masked or fully unmasked. For these cases, we can either generate a oob buffer load with no if condition, or we can generate a normal load with a if condition (if no fat_raw_buffer space). | 1 年前 | |
[mlir][amdgpu][rocdl] Add gfx1250 wmma ops (#165064) Update amdgpu.wmma op definition and implement amdgpu to rocdl conversion for new variants. | 9 个月前 |
| 文件 | 最后提交记录 | 最后更新时间 |
|---|---|---|
| 1 年前 | ||
| 11 个月前 | ||
| 1 年前 | ||
| 9 个月前 | ||
| 9 个月前 | ||
| 8 个月前 | ||
| 1 年前 | ||
| 9 个月前 |