AINewsnow

Feeling like I'm not getting enough out of my gpu?

So, for reference, my GPU is an RX 7900 XT, with 32 GB of RAM and a 7800X3D. I'm completely new to local LLMs, but I feel like I'm getting stiffed a little. I see people getting insane context windows while still maintaining great TPS. I'm currently using this model: https://huggingface.co/HauhauCS…

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-09-30 22:38 · r/LocalLLM
    Feeling like I'm not getting enough out of my gpu?

More stories

  1. NVIDIA Open Agent Safety Platform: A Reference for Continuous In-Silicon Agent Monitoring — NVIDIA Technical Blog
  2. FTC is investigating OpenAI, Anthropic and other AI companies over product risks — CNBC Technology
  3. OpenAI is sued over rogue AI Hugging Face cyberattack — CNBC Technology
  4. OpenAI hit with landmark lawsuit following Hugging Face hack — Axios AI+
  5. Nvidia announces security system to stop AI agents from going rogue — CBS News Technology
  6. OpenAI takes on rivals with launch of new dots agent — The Independent Tech
  7. Nova v3: 148M model, 1% of the data, same benchmark scores as SmolLM2-135M — r/huggingface
  8. Minimax H3 new comfy Model — r/StableDiffusion

Get the daily brief of stories like this at 6:30 every morning →