AINewsnow

Would you use a remote H3 text-encoder API so the 32B stays off your GPU?

H3 local users already know the pain: the DiT wants ~20–25GB, the encoder is a truncated Qwen3-VL-32B, and they do not co-fit on a 24–32GB card. Leave the encoder loaded and the sampler streams. Unload it and you wait on the next prompt. I’m considering a text-only encode API for T2VA: You send the…

Read the full story at r/comfyui ↗

Timeline · 1 report

  1. 2026-09-28 18:10 · r/comfyui
    Would you use a remote H3 text-encoder API so the 32B stays off your GPU?

More stories

  1. NVIDIA Open Agent Safety Platform: A Reference for Continuous In-Silicon Agent Monitoring — NVIDIA Technical Blog
  2. How we found 24 Android vulnerabilities using our open source AI security agent — GitHub Blog
  3. Introducing Claude Sonnet 5.5 on AWS — AWS Machine Learning Blog
  4. AMD to Buy Fei-Fei Li’s World Labs AI Startup for $8.2 Billion — Bloomberg AI
  5. Meta launches enterprise AI business seeking to cash in on vast spending — Financial Times AI
  6. Heads of OpenAI and Anthropic called to face Senate inquiry after rogue agent incidents — The Guardian AI
  7. Scoop: Anthropic's Dario Amodei to have White House dinner with Trump — Axios AI+
  8. OpenAI’s A.I. Went Rogue and Meddled With U.S. Government Websites — New York Times Technology

Get the daily brief of stories like this at 6:30 every morning →