Explaining Image Similarity with Automatically Extracted Concept Activation Vectors 文章

ArXiv CS.CV2026-07-31PAPERen作者: Isaac Roberts, Petra Bevandic, Alexander Schulz, Barbara Hammer

详细信息

来源站点
ArXiv CS.CV
作者
Isaac Roberts, Petra Bevandic, Alexander Schulz, Barbara Hammer
文章类型
PAPER
语言
en
发布日期
2026-07-31

摘要

arXiv:2607.28386v1 Announce Type: new Abstract: Image similarity underlies many computer vision applications, yet it is often unclear why two images receive a high or low similarity score. Existing explainability methods often rely on gradient-based attribution maps to provide local justifications for similarity. These approaches struggle to provide global insights into what specifically drives similarity in regions of an embedding space, such as texture, shape, or color. We introduce a model- and metric-agnostic framework that explains image similarity using Concept Activation Vectors (CAVs) extracted automatically via Sparse Autoencoders (SAEs). Given a pair of images, we perturb their embeddings along discovered concept directions and measure the resulting change in a chosen similarity function, yielding concept importances. For image pairs, we provide localization with concept attribution maps.