<?xml version="1.0" encoding="UTF-8"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Yestino - The Signal · Disable USE_FLASH_ATTENTION</title><link>https://yestino.com/entities/disable-use-flash-attention-db5bb4</link><description>Every event involving Disable USE_FLASH_ATTENTION</description><language>en</language><atom:link href="https://yestino.com/entities/disable-use-flash-attention-db5bb4/feed.xml" rel="self" type="application/rss+xml"/><item><title>trunk/68d20d4ee3956ceb5fecbc1676112aa32d2ad7d9: Disable USE_FLASH_ATTENTION when no CUDA arch supports it (#194502)</title><link>https://yestino.com/events/trunk-68d20d4ee3956ceb5fecbc1676112aa32d2ad7d9-disable-use-f-266804</link><guid isPermaLink="true">https://yestino.com/events/trunk-68d20d4ee3956ceb5fecbc1676112aa32d2ad7d9-disable-use-f-266804</guid><pubDate>Mon, 24 Aug 2026 02:06:01 GMT</pubDate><description>USE_FLASH_ATTENTION stays ON for any CUDA build regardless of target architecture, even though the kernels need sm80+ and can_use_flash_attention() rejects any…
Sources: PyTorch Releases</description></item></channel></rss>