Is anyone working on conversation compaction?
This story is from 2026-09-09. It is preserved in the archive; the latest stories are on the live feed.
In our chat app "harness" we recursively generate summaries, L1 → L2 → L3. L1 summaries are more factual extraction than coherence, then get rolled into a more storytelling L2. Then we keep a tail of always ~10 raw messages with timestamps. We're not coding so this keeps context really clean like 5…
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-09-09 09:30 · r/LocalLLaMA
Is anyone working on conversation compaction?