SnapBench: Benchmarking Snap-and-Ask Multimodal Retrieval for Mobile Interactions
- 类型:arxiv
- 标识:2608.29607
- 链接:https://arxiv.org/abs/2608.29607
- 主分类:rag
- 形态:benchmark
- TLDR:Mobile AI acts as a visual oracle, empowering users to snap a picture of something and ask for information. Snap-and-ask retrieval is now one of the most common entry points for mobile AI, yet photos are often blurry, while text questions may be short or mistyped. Existing benchmarks only test on clean inputs or do not isolate paired robustness in snap-and-ask retrieval. Therefore, we introduce SnapBench, the first paired benchmark for robust snap-and-ask multimodal retrieval, spanning 1,145 queries, 9,085 gallery items under 53 controlled corruption conditions with human annotations. We evalu
- 副分类:evaluation
- 待LLM分类:否
- 来源文件:
- /inbox/tom/_candidates/2026-09-03-agent-rag-longcontext-candidates.json