What parts of your agent stack do you actually run locally ... and what wasn’t worth it?
For agent builders, are you running embeddings, retrieval, reranking, memory, and logging on infrastructure you control? Or mixing local components with hosted APIs? I’m especially interested in hybrid setups, where you have local data and retrieval, with a hosted model for generation. How do you d…
Read the full story at r/AI_Agents ↗
Timeline · 1 report
- 2026-10-03 15:35 · r/AI_Agents
What parts of your agent stack do you actually run locally ... and what wasn’t worth it?