AINewsnow

I built a local way to export, search and continue chats from OpenRouter, LM Studio and AI Studio with llama.cpp or OpenRouter

This story is from 2026-09-13. It is preserved in the archive; the latest stories are on the live feed.

I had a lot of chats in OpenRouter across different models, with basically no proper way to bulk export or search them. So I wrote a js scraper for that. LM Studio was easier because chats are local, while Google AI Studio had the ugliest format. That became ThreadShelf: one local archive with sema…

Read the full story at r/LocalLLaMA ↗

Timeline · 1 report

  1. 2026-09-13 12:41 · r/LocalLLaMA
    I built a local way to export, search and continue chats from OpenRouter, LM Studio and AI Studio with llama.cpp or OpenRouter

More stories

  1. Google's Gemini AI hacks three other companies during security test — Sky News Technology
  2. Qwen3.8-Flash-Next-Heretic2-IQ4XS on Halogen Flash Server vs llama-server on Strix Halo: 2.3-7.7x prefill speedup with half the VRAM (+ vision works on BYO GGUF) — r/LocalLLM
  3. M2 Mac ultra128gb Qwen flash next — r/LocalLLM
  4. Multi-hour llama.cpp optimization experiments on Qwen MoE models, patches, benchmarks, and reproduction guides — r/LocalLLM
  5. Google’s Gemini AI hacked into other companies, adding to ‘rogue’ AI incidents — Washington Post AI
  6. M1 Max 32GB, trying to run Qwen 3.8 27B at decent speeds and context — r/LocalLLaMA
  7. I turned an asymetric pair of Tesla V100s PCIe both (16 GB + 32 GB) into a surprisingly capable local LLM lab — 1.38k prompt tok/s, 40 decode tok/s with qwen3.8 27B Q6 and Q8... — r/LocalLLaMA
  8. [Guide / Weights] Qwen 3.8 27B on Intel Arc: Why IQ quants crawl at 8 tok/s, why Q4_K outpaces sub-4bpw on Battlemage, and clean RCO GGUFs (16GB & 24GB) — r/LocalLLM

Get the daily brief of stories like this at 6:30 every morning →