trunk/4202fb82cb815b48e5ab34e8496352507160b6bc: [cuBLAS] Always eagerly allocate cuBLAS(Lt) workspaces (#194311)
- PyTorch: 511 events in the last 90 days
- PyTorch: 506th Release in the last 90 days
- Previous: earlier the same day · trunk/9a32e12fc5da06217fb6f90d6f3438eb609afd21: [dynamo] Guard autocast objects by value rather than identity (#194754)
What happened
authored with codex as discussed w/ @eellison , @ngimel ,~~~ just stashing this prototype here as performance doesn't look great on the hot path:~~~ AI-generated benchmark summary, provided for human review > > Benchmarked cached versus operation-scoped eager cuBLAS workspaces on an NVIDIA GB300, CC 10.3, CUDA 13.4. PyTorch was built for CC 10.0. Measurements are medians across five alternating cached/eager process…
Summary assembled by rule from the sources below