A dataset of word recognition accuracy, times, and prevalence for 4,562 verbs and 4,562 pseudoverbs of Spanish
摘要
The current database comprises behavioral data from 267 healthy Spanish adults aged 18 to 51, who participated in a visual lexical decision task conducted in a laboratory setting. Raw data on response accuracy and latencies are available for 4,565 verbs and 4,565 pseudoverbs, along with word prevalence—calculated from accuracy in that task. Additionally, a filtered dataset containing response times (RTs) for only correct responses, applying commonly used thresholds for outliers, is provided. Reliability analyses (via ICC) demonstrated good to excellent scores for the entire dataset. Criterion and construct validity, including a comparative analysis with other databases, were also examined, yielding satisfactory results. The new data can facilitate researchers in conducting virtual or pilot experiments, exploring novel research questions, or constructing models of word recognition. Furthermore, they may be of interest at the clinical level, such as in the selection of materials for cognitive rehabilitation, as the behavioral data and prevalence can be utilized to organize words based on their difficulty for training purposes.