| 文件 | 最后提交记录 | 最后更新时间 |
|---|---|---|
[MLIR][NVGPU] Fix nvgpu_arrive syntax in matmulBuilder.py (#113713) This patch updates the syntax for nvgpu_arrive Op in matmulBuilder.py. This fixes the compilation error for this test. For the warp-specialized matmul_kernel implementation, removing the WaitGroupSyncOp (after the mma-main-loop) fixes the hang observed. With these two fixes, the test compiles and executes successfully on an sm90a machine. Signed-off-by: Durgadoss R <durgadossr@nvidia.com> | 1 年前 | |
[mlir] GEMM Hopper Tensor Core Integration Test (#81478) | 2 年前 | |
[mlir] GEMM Hopper Tensor Core Integration Test (#81478) | 2 年前 |
| 文件 | 最后提交记录 | 最后更新时间 |
|---|---|---|
| 1 年前 | ||
| 2 年前 | ||
| 2 年前 |