2024
SINDER: Repairing the Singular Defects of DINOv2
ECCV 2024oral
"Vision Transformer models trained on large-scale datasets, although effective, often exhibit artifacts in the patch token they extract. While such defects can be alleviated by re-training the entire model with additional classification tokens, the underlying reasons for the presence of these tokens…