AINewsnow

Built a quick, sub-15ms Rust CLI/TUI to pack repos into prompts without burning 40k tokens on lockfiles and junk

Whenever I feed codebases into local models (Qwen, DeepSeek R1) or API models, the biggest annoyance is prompt pollution: - Lockfiles (`Cargo.lock`, `package-lock.json`) burning 30,000+ tokens for zero reason. - SVGs, binary files, or build artifacts slipping into the context. - Existing packers ta…

Read the full story at r/LocalLLaMA ↗

Timeline · 1 report

  1. 2026-10-04 17:18 · r/LocalLLaMA
    Built a quick, sub-15ms Rust CLI/TUI to pack repos into prompts without burning 40k tokens on lockfiles and junk

More stories

  1. Built a gateway so you can call DeepSeek, Qwen, Kimi, GLM, MiniMax with one key — USD billing, OpenAI-compatible — r/LocalLLM
  2. I put Gemma 4 26B and Qwen 3.8 27B (2x3090) in charge of a club in Championship Manager 97/98, against Claude, Grok and DeepSeek. It's running now. — r/LocalLLM
  3. best <40B alternatives to Qwen/Deepseek for (1) Coding (2) Long document QA test — r/LocalLLaMA
  4. If you’re not running local — do you use Chinese commercial LLMs (Qwen / GLM / MiniMax / etc)? — r/LocalLLM
  5. Who’s the current “king” of local LLMs for you — Qwen, Gemma, Llama, something else? — r/LocalLLM
  6. China's AI stack is collapsing search, commerce and payments into one loop — r/ArtificialInteligence
  7. Strata is seriously impressive, running Qwen 3.8 Flash Next on hermes at 512k context. — r/LocalLLM
  8. Qwen3.8-Flash-Next 177B running at 11–15 tok/s on a single RTX 5070 12GB + 32GB RAM DDR4 — r/LocalLLaMA

Get the daily brief of stories like this at 6:30 every morning →