Aligning Visual Contrastive learning models via Preference Optimization

Afzali, Amirabbas; Khodabandeh, Borna; Rasekh, Ali; JafariNodeh, Mahyar; kazemi, Sepehr; Gottschalk, Simon

Computer Science > Computer Vision and Pattern Recognition

arXiv:2411.08923 (cs)

[Submitted on 12 Nov 2024 (v1), last revised 26 Mar 2025 (this version, v3)]

Title:Aligning Visual Contrastive learning models via Preference Optimization

Authors:Amirabbas Afzali, Borna Khodabandeh, Ali Rasekh, Mahyar JafariNodeh, Sepehr kazemi, Simon Gottschalk

View PDF HTML (experimental)

Abstract:Contrastive learning models have demonstrated impressive abilities to capture semantic similarities by aligning representations in the embedding space. However, their performance can be limited by the quality of the training data and its inherent biases. While Preference Optimization (PO) methods such as Reinforcement Learning from Human Feedback (RLHF) and Direct Preference Optimization (DPO) have been applied to align generative models with human preferences, their use in contrastive learning has yet to be explored. This paper introduces a novel method for training contrastive learning models using different PO methods to break down complex concepts. Our method systematically aligns model behavior with desired preferences, enhancing performance on the targeted task. In particular, we focus on enhancing model robustness against typographic attacks and inductive biases, commonly seen in contrastive vision-language models like CLIP. Our experiments demonstrate that models trained using PO outperform standard contrastive learning techniques while retaining their ability to handle adversarial challenges and maintain accuracy on other downstream tasks. This makes our method well-suited for tasks requiring fairness, robustness, and alignment with specific preferences. We evaluate our method for tackling typographic attacks on images and explore its ability to disentangle gender concepts and mitigate gender bias, showcasing the versatility of our approach.

Subjects:	Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG)
Cite as:	arXiv:2411.08923 [cs.CV]
	(or arXiv:2411.08923v3 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2411.08923

Submission history

From: Amirabbas Afzali [view email]
[v1] Tue, 12 Nov 2024 08:14:54 UTC (1,597 KB)
[v2] Wed, 22 Jan 2025 23:58:03 UTC (3,200 KB)
[v3] Wed, 26 Mar 2025 11:37:00 UTC (5,321 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Aligning Visual Contrastive learning models via Preference Optimization

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Aligning Visual Contrastive learning models via Preference Optimization

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators