trunk/bf06d76b78b5735e759f6f2f4668c4baca964eee: [inductor] Make CUDA graph autotuning measurements consistent (#194649)
- PyTorch 近 90 天出现 1839 次
- PyTorch 近 90 天第 1815 次发版
- 上一次:同一天稍早 · ciflow/trunk/199479: Enable bugprone-unused-local-non-trivial-variable and remove hits
发生了什么
Human commentary: I want every candidate in an autotune decision measured under the same launch policy so the selected kernel reflects production execution. AI-assisted content: Problem Fixes #198171 Autotune requests can be created under one Inductor configuration and executed later in another process. Without carrying the effective CUDA-graph policy on the request, candidates from the same decision can be timed un…
摘要按规则整理自下方来源原文