Words are all you need? Language as an approximation for human similarity judgments

Marjieh, Raja; van Rijn, Pol; Sucholutsky, Ilia; Sumers, Theodore R.; Lee, Harin; Griffiths, Thomas L.; Jacoby, Nori

Computer Science > Computation and Language

arXiv:2206.04105 (cs)

[Submitted on 8 Jun 2022 (v1), last revised 23 Feb 2023 (this version, v3)]

Title:Words are all you need? Language as an approximation for human similarity judgments

Authors:Raja Marjieh, Pol van Rijn, Ilia Sucholutsky, Theodore R. Sumers, Harin Lee, Thomas L. Griffiths, Nori Jacoby

View PDF

Abstract:Human similarity judgments are a powerful supervision signal for machine learning applications based on techniques such as contrastive learning, information retrieval, and model alignment, but classical methods for collecting human similarity judgments are too expensive to be used at scale. Recent methods propose using pre-trained deep neural networks (DNNs) to approximate human similarity, but pre-trained DNNs may not be available for certain domains (e.g., medical images, low-resource languages) and their performance in approximating human similarity has not been extensively tested. We conducted an evaluation of 611 pre-trained models across three domains -- images, audio, video -- and found that there is a large gap in performance between human similarity judgments and pre-trained DNNs. To address this gap, we propose a new class of similarity approximation methods based on language. To collect the language data required by these new methods, we also developed and validated a novel adaptive tag collection pipeline. We find that our proposed language-based methods are significantly cheaper, in the number of human judgments, than classical methods, but still improve performance over the DNN-based methods. Finally, we also develop `stacked' methods that combine language embeddings with DNN embeddings, and find that these consistently provide the best approximations for human similarity across all three of our modalities. Based on the results of this comprehensive study, we provide a concise guide for researchers interested in collecting or approximating human similarity data. To accompany this guide, we also release all of the similarity and language data, a total of 206,339 human judgments, that we collected in our experiments, along with a detailed breakdown of all modeling results.

Comments:	Accepted to ICLR 2023, final revision. this https URL
Subjects:	Computation and Language (cs.CL); Machine Learning (cs.LG); Machine Learning (stat.ML)
Cite as:	arXiv:2206.04105 [cs.CL]
	(or arXiv:2206.04105v3 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2206.04105

Submission history

From: Raja Marjieh [view email]
[v1] Wed, 8 Jun 2022 18:09:19 UTC (30,194 KB)
[v2] Wed, 15 Jun 2022 17:31:17 UTC (11,153 KB)
[v3] Thu, 23 Feb 2023 18:44:23 UTC (4,063 KB)

Computer Science > Computation and Language

Title:Words are all you need? Language as an approximation for human similarity judgments

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Words are all you need? Language as an approximation for human similarity judgments

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators