Comparing human text classification performance and explainability with large language and machine learning models using eye-tracking

被引:0
|
作者
Venkatesh, Jeevithashree Divya [1 ]
Jaiswal, Aparajita [2 ]
Nanda, Gaurav [1 ]
机构
[1] Purdue Univ, Sch Engn Technol, W Lafayette, IN 47907 USA
[2] Purdue Univ, Ctr Intercultural Learning Mentorship Assessment &, W Lafayette, IN 47907 USA
来源
SCIENTIFIC REPORTS | 2024年 / 14卷 / 01期
关键词
Human-AI alignment; Large language models; Explainable AI; Eye tracking; Cognitive engineering; Human-computer interaction; MOVEMENTS; GAZE;
D O I
10.1038/s41598-024-65080-7
中图分类号
O [数理科学和化学]; P [天文学、地球科学]; Q [生物科学]; N [自然科学总论];
学科分类号
07 ; 0710 ; 09 ;
摘要
To understand the alignment between reasonings of humans and artificial intelligence (AI) models, this empirical study compared the human text classification performance and explainability with a traditional machine learning (ML) model and large language model (LLM). A domain-specific noisy textual dataset of 204 injury narratives had to be classified into 6 cause-of-injury codes. The narratives varied in terms of complexity and ease of categorization based on the distinctive nature of cause-of-injury code. The user study involved 51 participants whose eye-tracking data was recorded while they performed the text classification task. While the ML model was trained on 120,000 pre-labelled injury narratives, LLM and humans did not receive any specialized training. The explainability of different approaches was compared based on the top words they used for making classification decision. These words were identified using eye-tracking for humans, explainable AI approach LIME for ML model, and prompts for LLM. The classification performance of ML model was observed to be relatively better than zero-shot LLM and non-expert humans, overall, and particularly for narratives with high complexity and difficult categorization. The top-3 predictive words used by ML and LLM for classification agreed with humans to a greater extent as compared to later predictive words.
引用
收藏
页数:12
相关论文
共 50 条
  • [41] Using eye-tracking and EEG to study the mental processing demands during learning of text-picture combinations
    Scharinger, Christian
    Schueler, Anne
    Gerjets, Peter
    INTERNATIONAL JOURNAL OF PSYCHOPHYSIOLOGY, 2020, 158 : 201 - 214
  • [42] Text Classification by CEFR Levels Using Machine Learning Methods and the BERT Language Model
    Lagutina, N. S.
    Lagutina, K. V.
    Brederman, A. M.
    Kasatkina, N. N.
    AUTOMATIC CONTROL AND COMPUTER SCIENCES, 2024, 58 (07) : 869 - 878
  • [43] Large language models in orthopedics: An exploratory research trend analysis and machine learning classification
    Garcia, Ausberto Velasquez
    Minami, Masataka
    Mejia-Rodriguez, Manuel
    Ortiz-Morales, Jorge Rolando
    Radice, Fernando
    JOURNAL OF ORTHOPAEDICS, 2025, 66 : 110 - 118
  • [44] Comparing the Visual Perception According to the Performance Using the Eye-Tracking Technology in High-Fidelity Simulation Settings
    Tanoubi, Issam
    Tourangeau, Mathieu
    Sodoke, Komi
    Perron, Roger
    Drolet, Pierre
    Belanger, Marie-Eve
    Morris, Judy
    Ranger, Caroline
    Paradis, Marie-Rose
    Robitaille, Arnaud
    Georgescu, Mihai
    BEHAVIORAL SCIENCES, 2021, 11 (03)
  • [45] Using Large Language Models to Translate Machine Results to Human Results
    Niraula, Trishna
    Stubblefield, Jonathan
    14TH ACM CONFERENCE ON BIOINFORMATICS, COMPUTATIONAL BIOLOGY, AND HEALTH INFORMATICS, BCB 2023, 2023,
  • [46] Human-interpretable clustering of short text using large language models
    Miller, Justin K.
    Alexander, Tristram J.
    ROYAL SOCIETY OPEN SCIENCE, 2025, 12 (01):
  • [47] Development and Comparison of Multiple Emotion Classification Models in Indonesia Text Using Machine Learning
    Zamsuri, Ahmad
    Defit, Sarjon
    Nurcahyo, Gunadi Widi
    JOURNAL OF ADVANCES IN INFORMATION TECHNOLOGY, 2024, 15 (04) : 519 - 531
  • [48] Transfer effects and word-class-dependent improvement of foreign language text comprehension: An empirical study using eye-tracking
    Hiller, Laura Nathalie
    Futner, Marco
    Sachse, Pierre
    Martini, Markus
    INTERNATIONAL JOURNAL OF PSYCHOLOGY, 2012, 47 : 375 - 375
  • [49] Automated object detection in mobile eye-tracking research: comparing manual coding with tag detection, shape detection, matching, and machine learning
    Segijn, Claire M.
    Menheer, Pernu
    Lee, Garim
    Kim, Eunah
    Olsen, David
    Mohr, Alicia Hofelich
    COMMUNICATION METHODS AND MEASURES, 2024,
  • [50] Identification of Scientific Texts Generated by Large Language Models Using Machine Learning
    Soto-Osorio, David
    Sidorov, Grigori
    Chanona-Hernandez, Liliana
    Lopez-Ramirez, Blanca Cecilia
    COMPUTERS, 2024, 13 (12)