Evidence map›Paper›PMID 42711454›Full record

ArticleBehavior research methods2026

Surprisal from Large Language Models as an individual difference: Capturing the individual regardless of the text answer.

José Ángel Martínez-Huertas, Alejandro Martínez-Mingo, Antonio López Rosell, Rodrigo S Kreitchmann, Guillermo Jorge-Botana

Abstract read
PubMed Publisher
In one paragraph

Article in Behavior research methods, 2026. The graph could read no effect estimate from its abstract, so it casts no vote on the map. Not yet cited in PubMed.

0numbers the graph read from it
0cells of the map it votes in
0citing papers in PubMed
–field-weighted citation impact
1 · What the graph read from it

What it found

Each row is one number read from the abstract, on the scale the paper reported it, with its interval. Left of the dashed line favours the treatment, right favours the comparator. Under each row is the sentence it came from. New to these charts? A ten-minute tutorial.

The abstract states no effect estimate the extractor could read, or names no intervention and outcome on the map, so this paper lights no cell and moves no belief. It is still indexed, cited and linked below.

2 · The registry

The trial behind it

Trials whose registry record cites this paper, or whose number appears in the abstract. A trial that started after this paper was published is citing it as background, not reporting it.

Neither the registry nor the abstract names a trial number. If this is a trial report, that itself is worth knowing.

3 · Its place in the literature

Who cites it

0 citing papers in PubMed.

No citing paper in PubMed yet.

4 · The record

Corrections and comments

PubMed lists nothing against this paper. Absence here is not a guarantee, only a check that was made.

5 · Who and what money

Authors and funding

5 authors.

José Ángel Martínez-HuertasDepartment of Methodology of Behavioral Sciences, National Distance Education University, Madrid, Spain. jamartinez@psi.uned.es.ORCID http://orcid.org/0000-0002-6700-6832
Alejandro Martínez-MingoDepartment of Methodology of Behavioral Sciences, National Distance Education University, Madrid, Spain.
Antonio López RosellDepartment of Methodology of Behavioral Sciences, National Distance Education University, Madrid, Spain.
Rodrigo S KreitchmannDepartment of Methodology of Behavioral Sciences, National Distance Education University, Madrid, Spain.
Guillermo Jorge-BotanaDepartment of Psychobiology and Behavioral Sciences Methods, Complutense University of Madrid, Madrid, Spain.

Funding

No grant is acknowledged in the PubMed record.

6 · The paper itself

Abstract

Surprisal is a measure from information theory that quantifies how unexpected an event is within the context of a particular system. In Large Language Models, higher surprisal of particular words indicates lower predictability or greater difficulty for cognitive processing. In language production, surprisal can be understood as a quantification of individual choices plus other produced linguistic characteristics such as selected content. Thus, this measure could be a quantifiable characteristic of how individuals produce their discourse and its predictability. However, it is not clear whether these measures can be considered an actual stable characteristic of the individuals or a measure of particular text answers without generalization to other language production of the same individual (i.e., a stable individual difference across language-based items). In a series of four studies, we analyzed the stability of surprisal measures across different language-based items of different tasks taken from previous publications. The reliability of mean surprisal scores was found to be excellent in the language-based task consisting of different language-based items, and to vary within the items themselves. We showed that surprisal is a stable individual characteristic regardless of the item but is significantly influenced by the length of the text answer. We also analyzed the relationship between mean surprisal scores and relevant variables ranging from language skills or domain-general cognitive skills to personality traits across the different datasets. We concluded that mean surprisal scores are related to certain cognitive functions and may also reflect alternative styles of speech or communication in speech production.

Indexed as

IndividualityLanguageHumansLarge Language ModelsIndividual differencesLanguageLarge language modelsNeural networksPsychological constructSurprisal

Identifiers

PMID42711454

What OpenQuestion holds

Textmetadata
Read underepoch 390

Registered trials

None linked

Read under generation 80e0d062 · epoch 390. Bibliography from PubMed, PubMed Central and OpenAlex; grants from NIH RePORTER; trial links from ClinicalTrials.gov; estimates, votes and beliefs from the OpenQuestion graph.