Exploring the psychology of LLMs’ moral and legal reasoning

GUILHERME DA FRANCA COUTO FERNANDES DE ALMEIDA; Nunes, José Luiz; Engelmann, Neele; Wiegmann, Alex; Araújo, Marcelo de

Exploring the psychology of LLMs’ moral and legal reasoning

Autores

GUILHERME DA FRANCA COUTO FERNANDES DE ALMEIDA

Nunes, José Luiz

Engelmann, Neele

Wiegmann, Alex

Araújo, Marcelo de

Tipo de documento

Artigo Científico

Data

2024

Arquivos

Primário Acesso_Primeira Pagina_Exploring the psychology of LLMs’ moral and legal reasoning.pdf (237.32 KB)

ACESSO_RESTRITO_Artigo_2024_Exploring_the_psychology_of_LLMs_moral_and_legal_reasoning_TC.pdf (2.17 MB)

Resumo

Large language models (LLMs) exhibit expert-level performance in tasks across a wide range of different domains. Ethical issues raised by LLMs and the need to align future versions makes it important to know how state of the art models reason about moral and legal issues. In this paper, we employ the methods of experimental psychology to probe into this question. We replicate eight studies from the experimental literature with instances of Google's Gemini Pro, Anthropic's Claude 2.1, OpenAI's GPT-4, and Meta's Llama 2 Chat 70b. We find that alignment with human responses shifts from one experiment to another, and that models differ amongst themselves as to their overall alignment, with GPT-4 taking a clear lead over all other models we tested. Nonetheless, even when LLM-generated responses are highly correlated to human responses, there are still systematic differences, with a tendency for models to exaggerate effects that are present among humans, in part by reducing variance. This recommends caution with regards to proposals of replacing human participants with current state-of-the-art LLMs in psychological research and highlights the need for further research about the distinctive aspects of machine psychology

Palavras-chave

AI Ethics; Experimental jurisprudence; Ethics of artificial intelligence; Machine Behavior; Moral psychology; Machine psychology; Large language models

Titulo de periódico

Artificial Intelligence

Texto completo

https://www.sciencedirect.com/science/article/pii/S000437022400081X?via%3Dihub

DOI

Idioma

en

URI

https://repositorio.insper.edu.br/handle/11224/6965

Área do Conhecimento CNPQ

CIENCIAS SOCIAIS APLICADAS

Coleções

Coleção de Artigos Acadêmicos

Página do item completo