viable/strict/1790773504: Extend `native_layer_norm` param dtype check from CUDA to XPU (#198647)
- PyTorch 近 90 天出现 1696 次
- PyTorch 近 90 天第 1675 次发版
- 上一次:同一天稍早 · trunk/071350f5555440e78fd85b72674f5b8ebcc42295
发生了什么
_check_native_layer_norm_cuda_param_dtype early-returned on non-CUDA devices, so the decomposition skipped the weight/bias dtype validation on XPU and diverged from eager under aot_eager_decomp_partition . XPU has identical kernel semantics (including the num_rows == 0 carve-out), so gate on both devices and rename the helper accordingly. Unskips test_layer_norm_mixed_dtype_aot_eager_decomp_partition_errors for XPU.…
摘要按规则整理自下方来源原文