Suggest architecture/pipeline for general object detection + VLM call afterwards
This story is from 2026-09-09. It is preserved in the archive; the latest stories are on the live feed.
Hi, I am looking for the following model selection/inference pipeline. Goal is something like this: 1) detect human -> describe human 2) detect human -> detect objects in human hand -> describe objects 3) detect animal -> get specific animal type 4) detect general object (i.e package) --- What are…
Read the full story at r/computervision ↗
Timeline · 1 report
- 2026-09-09 05:46 · r/computervision
Suggest architecture/pipeline for general object detection + VLM call afterwards