Return
LeCoder: A large-scale automated coder for coding errors in word-production tasks
DOI:10.3758/s13428-026-02948-8.png)
Abstract
En 中文
Speech errors have been instrumental in advancing our understanding of the architecture of the language production system, the nature of its representations, and its disorders. To be most informative, researchers usually need large amounts of data. Hand-coding such data can be both cumbersome and subjective. This paper presents LeCoder, the first open-source, automated error coder for English word and naming data, which uses a data-driven approach grounded in large-scale corpora to quantify the target–response relationship, allowing it to be flexible, scalable, and generalizable across new datasets. By testing the coder on two datasets from two aphasia labs that have been carefully coded by trained research assistants, we first establish that LeCoder has high accuracy when compared to expert coders, and in certain cases, offers a more logical categorization than human coders. We then show, using robust machine-learning approaches, that LeCoder’s performance generalizes to new participants and items it has never encountered before. Collectively, these findings encourage the use of LeCoder across labs for more objective coding of speech errors, which will, in turn, increase replicability of findings in all subfields of research that use speech error analysis, including neuropsychological research.
Keywords:
Speech errors
Automated coding
Semantic similarity
Phonological similarity
Aphasia
AI Summary
Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.
Journal
IF:
3.9
Papers:
720
Citations:
3.6W
Organization
Cited Papers
Similarity-induced interference or facilitation in language production reflects representation, not selection
Cognition
IF0
Understanding semantic and phonological processing deficits in adults with aphasia: effects of category and typicality
Aphasiology
IF0

