Skip to content
X · @ylecun · X / Twitter

RT Lukas Kuhn: 🔥 We introduce LeVLJEPA: the first fully non-contrastive end-to-end vision-language pretraining method competitive with CLIP & SigLI…

RT Lukas Kuhn🔥 We introduce LeVLJEPA: the first fully non-contrastive end-to-end vision-language pretraining method competitive with CLIP & SigLIP 💪🏼👀 No negatives. No temperature. No momentum encoder. No teacher-student.TL;DR: LeVLJEPA learns image to text structure by prediction: each modality predicts the other's em