Hybrid Human-LLM Corpus Construction and LLM Evaluation for Rare Linguistic Phenomena

Weissweiler, Leonie; Köksal, Abdullatif; Schütze, Hinrich

Computer Science > Computation and Language

arXiv:2403.06965 (cs)

[Submitted on 11 Mar 2024]

Title:Hybrid Human-LLM Corpus Construction and LLM Evaluation for Rare Linguistic Phenomena

Authors:Leonie Weissweiler, Abdullatif Köksal, Hinrich Schütze

View PDF HTML (experimental)

Abstract:Argument Structure Constructions (ASCs) are one of the most well-studied construction groups, providing a unique opportunity to demonstrate the usefulness of Construction Grammar (CxG). For example, the caused-motion construction (CMC, ``She sneezed the foam off her cappuccino'') demonstrates that constructions must carry meaning, otherwise the fact that ``sneeze'' in this context causes movement cannot be explained. We form the hypothesis that this remains challenging even for state-of-the-art Large Language Models (LLMs), for which we devise a test based on substituting the verb with a prototypical motion verb. To be able to perform this test at statistically significant scale, in the absence of adequate CxG corpora, we develop a novel pipeline of NLP-assisted collection of linguistically annotated text. We show how dependency parsing and GPT-3.5 can be used to significantly reduce annotation cost and thus enable the annotation of rare phenomena at scale. We then evaluate GPT, Gemini, Llama2 and Mistral models for their understanding of the CMC using the newly collected corpus. We find that all models struggle with understanding the motion component that the CMC adds to a sentence.

Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:2403.06965 [cs.CL]
	(or arXiv:2403.06965v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2403.06965

Submission history

From: Leonie Weissweiler [view email]
[v1] Mon, 11 Mar 2024 17:47:47 UTC (8,806 KB)

Computer Science > Computation and Language

Title:Hybrid Human-LLM Corpus Construction and LLM Evaluation for Rare Linguistic Phenomena

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Hybrid Human-LLM Corpus Construction and LLM Evaluation for Rare Linguistic Phenomena

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators