How are you guys getting things done in low context window of 64K
I am running 5080 + 96GB RAM But how are you guys building real projects on low context window as these small models need more system prompt/skills/hand holding to get correct outputs compared to cloud models. I am using OMP
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-10-07 13:48 · r/LocalLLM
How are you guys getting things done in low context window of 64K