Understanding Your Agent: Leveraging Large Language Models for Behavior Explanation

Zhang, Xijia; Guo, Yue; Stepputtis, Simon; Sycara, Katia; Campbell, Joseph

Computer Science > Machine Learning

arXiv:2311.18062 (cs)

[Submitted on 29 Nov 2023]

Title:Understanding Your Agent: Leveraging Large Language Models for Behavior Explanation

Authors:Xijia Zhang, Yue Guo, Simon Stepputtis, Katia Sycara, Joseph Campbell

View PDF

Abstract:Intelligent agents such as robots are increasingly deployed in real-world, safety-critical settings. It is vital that these agents are able to explain the reasoning behind their decisions to human counterparts; however, their behavior is often produced by uninterpretable models such as deep neural networks. We propose an approach to generate natural language explanations for an agent's behavior based only on observations of states and actions, thus making our method independent from the underlying model's representation. For such models, we first learn a behavior representation and subsequently use it to produce plausible explanations with minimal hallucination while affording user interaction with a pre-trained large language model. We evaluate our method in a multi-agent search-and-rescue environment and demonstrate the effectiveness of our explanations for agents executing various behaviors. Through user studies and empirical experiments, we show that our approach generates explanations as helpful as those produced by a human domain expert while enabling beneficial interactions such as clarification and counterfactual queries.

Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI)
Cite as:	arXiv:2311.18062 [cs.LG]
	(or arXiv:2311.18062v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2311.18062

Submission history

From: Xijia Zhang [view email]
[v1] Wed, 29 Nov 2023 20:16:23 UTC (599 KB)

Computer Science > Machine Learning

Title:Understanding Your Agent: Leveraging Large Language Models for Behavior Explanation

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Understanding Your Agent: Leveraging Large Language Models for Behavior Explanation

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators