Paper page - SOCO: Benchmarking Semantic Object Correspondence in Vision Foundation Models
…and keypoint descriptions, revealing gaps between language-grounded localization and visual correspondence while demonstrating strong prediction of downstream task performance. Generated by Qwen/Qwen2.5-Coder-32B-Instruct Measuring structured object understanding…