Finally trying out Agentic stuff and now the question is how do I get my text in the real world in to my computer… OCR, Vision, etc.
I’ve come to realize I’ve captured a lot of text with my phone. Sure it has auto OCR on phone but by the time it’s on my computer it’s gone. I rather automate the process. Debating between vision and OCR or both with agentic harness voting on which is likely better. I currently plan to use pi.dev o…
Read the full story at r/LocalLLM ↗