AINewsnow

ExLlamaSharp v1.3.2: what shipped

This story is from 2026-09-08. It is preserved in the archive; the latest stories are on the live feed.

ExLlamaSharp v1.3.2 is out. Local LLM server for Windows with NVIDIA GPUs — OpenAI-compatible /v1 , Blazor admin, and EXL3 inference. Release notes Summary Fix false installation-failed dialog in desktop mode (Tray started after /health probe). Installer no longer aborts the Inno wizard when /healt…

Read the full story at DEV Community — AI ↗

Timeline · 1 report

  1. 2026-09-08 08:49 · DEV Community — AI
    ExLlamaSharp v1.3.2: what shipped

More stories

  1. Flyweight: open-source C++/CUDA engine for running MoE models bigger than your VRAM on one GPU + system RAM. First PyPI release, looking for contributors. — r/LocalLLaMA
  2. Elon Musk talks up AI safety while fighting regulation in wild week of strange alliances — CNBC Technology
  3. I built a small proxy that lets Claude Desktop / Claude Code run on local models and NVIDIA's free API, sharing it in case it's useful — r/LocalLLM
  4. Anthropic, OpenAI, SpaceXAI, Google sued over call to ‘pace’ AI development — Politico Technology
  5. Gemini Hacked Three Companies in First Known Breakout by Google’s AI — Wall Street Journal Technology
  6. NVIDIA CEO Jensen Huang rejects ‘AI will end the world’ claim, yet cautions ‘we should go as fast as we can but...’ — Mint AI
  7. Introducing the Australian Youth Safety Blueprint — OpenAI News
  8. Anthropic selects Accenture as first embedded evaluator to help implement Amodei's slowdown proposal — CNBC Technology

Get the daily brief of stories like this at 6:30 every morning →