AINewsnow

RDNA4 owners on ComfyUI: I made SageAttention actually fast on the 9070 XT, here's the build plus every caveat

This story is from 2026-10-01. It is preserved in the archive; the latest stories are on the live feed.

TL;DR: A drop-in sageattention build for RX 9070 / 9070 XT on Windows . The common case (fp16, head_dim 128) runs on an fp8 attention kernel I wrote by hand in HIP. Against the existing gfx12 port (SageAttention PR #368), each attention call is 1.07–1.21× faster non-causal and 1.72–1.96× faster cau…

Read the full story at r/StableDiffusion ↗

Timeline · 1 report

  1. 2026-10-01 23:08 · r/StableDiffusion
    RDNA4 owners on ComfyUI: I made SageAttention actually fast on the 9070 XT, here's the build plus every caveat

More stories

  1. NVIDIA DGX Spark 64GB Gives Developers More Ways to Build and Scale Local AI — NVIDIA Blog
  2. Trump’s big AI move: ‘Super Intelligence Force’ launched, Jay Clayton named AI czar — Mint AI
  3. A model guide for the GPT-6 family — OpenAI News
  4. An OpenAI safety employee has quit and is sounding the alarm — The Verge AI
  5. Introducing Oscilloscope Diffusion — r/comfyui
  6. OpenAI fires 3 AI safety researchers for allegedly sharing confidential company information — Mint AI
  7. Apple says it's tightening macOS Full Disk Access' controls due to new risks from AI agents — TechCrunch AI
  8. Google launches satellite to test feasibility of building data centers in space — NPR Technology

Get the daily brief of stories like this at 6:30 every morning →