AINewsnow

ExLlamaSharp v1.3.2.1: what shipped

This story is from 2026-09-08. It is preserved in the archive; the latest stories are on the live feed.

ExLlamaSharp v1.3.2.1 is out. Local LLM server for Windows with NVIDIA GPUs — OpenAI-compatible /v1 , Blazor admin, and EXL3 inference. Release notes Summary - LAN access toggle in Settings/Setup with firewall rule helper (Tray elevated netsh). - Worker hardening: prompt_too_long / kv_cache_full fa…

Read the full story at DEV Community — AI ↗

Timeline · 1 report

  1. 2026-09-08 15:06 · DEV Community — AI
    ExLlamaSharp v1.3.2.1: what shipped

More stories

  1. Amid growing AI fears, King Charles meets with industry leaders in Scotland — NPR Technology
  2. Flyweight: open-source C++/CUDA engine for running MoE models bigger than your VRAM on one GPU + system RAM. First PyPI release, looking for contributors. — r/LocalLLaMA
  3. Huang and Zuckerberg back AI safety without a slowdown: Why Big Tech is betting on market forces to police AI — Mint AI
  4. Elon Musk talks up AI safety while fighting regulation in wild week of strange alliances — CNBC Technology
  5. The AI Slowdown Debate Crashed Salesforce’s Party — Wired AI
  6. King Charles warns of 'existential danger' of AI falling into wrong hands — BBC Technology
  7. King Charles to press Nvidia, OpenAI, Anthropic leaders on AI safety at summit — CNBC Technology
  8. Would you buy branded clothing from your favourite tech firm? — BBC Technology

Get the daily brief of stories like this at 6:30 every morning →