viable/strict/1790977859: [ROCm] Enable test_grouped_mm on ROCm (#199048)
- PyTorch 近 90 天出现 1842 次
- PyTorch 近 90 天第 1818 次发版
- 上一次:同一天稍早 · viable/strict/1790976466
发生了什么
DistMatrixOpsTest.test_grouped_mm was skipped on ROCm via @unittest.skipIf(TEST_WITH_ROCM, "ROCm doesn't support CUTLASS"). That reason is a misnomer for this test: it is pure bf16, and on ROCm the backend="cutlass" parametrization routes to the arch-independent _grouped_mm fallback (CUTLASS is never invoked), while the backend="cublaslt" variants self-skip via an existing CUDA-Toolkit-version guard. Remove the ROCm…
摘要按规则整理自下方来源原文