AINewsnow

Twinny: Self-Hosted, Privacy-First Copilot for VS Code Without Cloud Lock-In

This story is from 2026-10-10. It is preserved in the archive; the latest stories are on the live feed.

Twinny: Self-Hosted, Privacy-First Copilot for VS Code Without Cloud Lock-In TL;DR twinny is an open-source, fully private AI coding companion for Visual Studio Code that bridges the gap between your editor and local LLM runners like Ollama, LM Studio, or llama.cpp. By decoupling AI completions and…

Read the full story at DEV Community — AI ↗

Timeline · 1 report

  1. 2026-10-10 06:29 · DEV Community — AI
    Twinny: Self-Hosted, Privacy-First Copilot for VS Code Without Cloud Lock-In

More stories

  1. Qwen 3.6 35B A3B: 131K context + vision on 6GB VRAM — r/LocalLLaMA
  2. feat: add GLM5Next MTP, optimize by pwilkin · Pull Request #29928 · ggml-org/llama.cpp — r/LocalLLaMA
  3. Tested Mellum2.1-12B-A2.5B on PI Coding Agent - surprisingly usable, but not great at one-shot projects — r/LocalLLaMA
  4. Running the uncensored Qwen3.8-27B (HauhauCS) on a 4090 at 262K context and ~130 tok/s — r/LocalLLaMA
  5. Java vllm-like framwork claims 90% of perfomance of llama.cpp on local inference on NVIDIA GPUs by compiling Java to CUDA and cuTile — r/LocalLLM
  6. Found a fix for AMD RX 6600 100% CPU usage — r/LocalLLM
  7. I made a free Mac app that runs local models: chat, pictures, video and voices — r/LocalLLM
  8. I have an ESC4000 G3 with 8x T4s in it - what is the fastest way I can deploy Qwen3.5-9B for about 10-15 users concurrently: currently using llama.cpp — r/LocalLLaMA

Get the daily brief of stories like this at 6:30 every morning →