arXiv (All)open access
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval
information retrieval, artificial intelligence, computation and language, computer vision and pattern recognition
This document is indexed with metadata only — full text is not available in the archive for this record.
Open the official source ↗
Related documents
Record · ID 974486
Retrieved via
Conceptio — every document is proof-bundled with source, license, and retrieval metadata.