← Back to events
ActiveAIRelease

viable/strict/1790977859: [ROCm] Enable test_grouped_mm on ROCm (#199048)

Photo: PyTorch Releases

What happened

DistMatrixOpsTest.test_grouped_mm was skipped on ROCm via @unittest.skipIf(TEST_WITH_ROCM, "ROCm doesn't support CUTLASS"). That reason is a misnomer for this test: it is pure bf16, and on ROCm the backend="cutlass" parametrization routes to the arch-independent _grouped_mm fallback (CUTLASS is never invoked), while the backend="cublaslt" variants self-skip via an existing CUDA-Toolkit-version guard. Remove the ROCm…

Summary assembled by rule from the sources below

Why it's spreading

Sources