PixelUp: Zero-Shot Semantic Feature Upsampling for Fine-Grained Vision Tasks
arXiv:2608.02792v1 Announce Type: new Abstract: Self-supervised Vision Foundation Models (VFMs) have become essential backbones for downstream tasks due to their strong and transferable visual representations. However,…