← 返回事件
持续讨论AI发版

viable/strict/1790725852: [inductor] Partition cross-device fallbacks from CUDA graphs (#190555)

图:PyTorch Releases

发生了什么

Summary Split any extern kernel whose FX node reads or produces tensors on more than one device, ignoring meta tensors. This covers FallbackKernel , ExternKernelOut (custom ops with a Tag.out overload), IndexPutFallback , and multi-output ops together with their MultiOutput children. Keep same-device kernels eligible for CUDA graph capture. Rename and simplify the deterministic device_put regression. Add regressions…

摘要按规则整理自下方来源原文

为什么在扩散

来源