Analyzing Local Representations of Self-supervised Vision Transformers | allainews.com

March 22, 2024, 4:46 a.m. | Ani Vanyan, Alvard Barseghyan, Hakob Tamazyan, Vahan Huroyan, Hrant Khachatrian, Martin Danelljan

cs.CV updates on arXiv.org arxiv.org

arXiv:2401.00463v2 Announce Type: replace
Abstract: In this paper, we present a comparative analysis of various self-supervised Vision Transformers (ViTs), focusing on their local representative power. Inspired by large language models, we examine the abilities of ViTs to perform various computer vision tasks with little to no fine-tuning. We design evaluation framework to analyze the quality of local, i.e.\ patch-level, representations in the context of few-shot semantic segmentation, instance identification, object retrieval and tracking. We discover that contrastive learning based methods …

abstract analysis analyze arxiv comparative analysis computer computer vision cs.cv design evaluation fine-tuning framework language language models large language large language models paper power tasks transformers type vision vision transformers

More from arxiv.org / cs.CV updates on arXiv.org

3D Human Pose Perception from Egocentric Stereo Videos 19 hours ago | arxiv.org

abstract arxiv compact cs.cv +11

SqueezeSAM: User friendly mobile interactive segmentation 19 hours ago | arxiv.org

abstract architecture arxiv computational +21

Polarimetric Light Transport Analysis for Specular Inter-reflection 19 hours ago | arxiv.org

abstract analysis arxiv cs.cv +13

Density-Guided Dense Pseudo Label Selection For Semi-supervised Oriented Object Detection 19 hours ago | arxiv.org

arxiv cs.cv detection object +4

CtxMIM: Context-Enhanced Masked Image Modeling for Remote Sensing Image Understanding 19 hours ago | arxiv.org

abstract arxiv clear context +19

nnSAM: Plug-and-play Segment Anything Model Improves nnUNet Performance 19 hours ago | arxiv.org

abstract arxiv clinical cs.cv +21

CoFiI2P: Coarse-to-Fine Correspondences for Image-to-Point Cloud Registration 19 hours ago | arxiv.org

arxiv cloud cs.ai cs.cv +5

Detail Reinforcement Diffusion Model: Augmentation Fine-Grained Visual Categorization in Few-Shot Conditions 19 hours ago | arxiv.org

abstract annotated data arxiv augmentation +18

Generative Image Dynamics 19 hours ago | arxiv.org

abstract arxiv clothes collection +15

Software Engineer for AI Training Data (School Specific)

@ G2i Inc | Remote

View on ai-jobs.net

Software Engineer for AI Training Data (Python)

@ G2i Inc | Remote

View on ai-jobs.net

Software Engineer for AI Training Data (Tier 2)

@ G2i Inc | Remote

View on ai-jobs.net

Data Engineer

@ Lemon.io | Remote: Europe, LATAM, Canada, UK, Asia, Oceania

View on ai-jobs.net

Artificial Intelligence – Bioinformatic Expert

@ University of Texas Medical Branch | Galveston, TX

View on ai-jobs.net

Lead Developer (AI)

@ Cere Network | San Francisco, US

View on ai-jobs.net