HiPerViT: A Hierarchical Perceiver-Vision Transformer Architecture for Multi-Scale Texture Recognition
This story is from 2026-09-11. It is preserved in the archive; the latest stories are on the live feed.
arXiv:2609.10917v1 Announce Type: new Abstract: Texture recognition remains challenging for modern vision models because discriminative evidence is often carried by higher-order spatial statistics rather than by object shape alone. While Vision Transformers provide strong long-range modeling capaci…
Read the full story at arXiv cs.CV ↗
Timeline · 1 report
- 2026-09-11 04:00 · arXiv cs.CV
HiPerViT: A Hierarchical Perceiver-Vision Transformer Architecture for Multi-Scale Texture Recognition