Understanding Medical Conversations: Rich Transcription, Confidence Scores & Information Extraction

被引:0
|
作者
Soltau, Hagen [1 ]
Wang, Mingqiu [1 ]
Shafran, Izhak [1 ]
El Shafey, Laurent [1 ]
机构
[1] Google, Mountain View, CA 94043 USA
来源
INTERSPEECH 2021 | 2021年
关键词
SPEECH RECOGNITION;
D O I
暂无
中图分类号
R36 [病理学]; R76 [耳鼻咽喉科学];
学科分类号
100104 ; 100213 ;
摘要
In this paper, we describe novel components for extracting clinically relevant information from medical conversations which will be available as Google APIs. We describe a transformerbased Recurrent Neural Network Transducer (RNN-T) model tailored for long-form audio, which can produce rich transcriptions including speaker segmentation, speaker role labeling, punctuation and capitalization. On a representative test set, we compare performance of RNN-T models with different encoders, units and streaming constraints. Our transformer-based streaming model performs at about 20% WER on the ASR task, 6% WDER on the diarization task, 43% SER on periods, 52% SER on commas, 43% SER on question marks and 30% SER on capitalization. Our recognizer is paired with a confidence model that utilizes both acoustic and lexical features from the recognizer. The model performs at about 0.37 NCE. Finally, we describe a RNN-T based tagging model. The performance of the model depends on the ontologies, with F-scores of 0.90 for medications, 0.76 for symptoms, 0.75 for conditions, 0.76 for diagnosis, and 0.61 for treatments. While there is still room for improvement, our results suggest that these models are sufficiently accurate for practical applications.
引用
收藏
页码:4418 / 4422
页数:5
相关论文
共 50 条
  • [11] TEACHING MEDICAL STUDENTS HOW TO HAVE BRAVE CONVERSATIONS TO PROMOTE UNDERSTANDING
    Nwankwo, O.
    Ongyiu, A.
    Sajdak, G.
    Hayton, A.
    ChenFeng, J.
    JOURNAL OF INVESTIGATIVE MEDICINE, 2024, 72 (01) : 309 - 309
  • [12] Architecture of a medical information extraction system
    Bekhouche, D
    Pollet, Y
    Grilheres, B
    Denis, X
    NATURAL LANGUAGE PROCESSING AND INFORMATION SYSTEMS, 2004, 3136 : 380 - 387
  • [13] A Span Extraction Approach for Information Extraction on Visually-Rich Documents
    Nguyen, Tuan-Anh D.
    Vu, Hieu M.
    Nguyen Hong Son
    Minh-Tien Nguyen
    DOCUMENT ANALYSIS AND RECOGNITION, ICDAR 2021, PT II, 2021, 12917 : 353 - 363
  • [14] Understanding Medical Conversations with Scattered Keyword Attention and Weak Supervision from Responses
    Shi, Xiaoming
    Hu, Haifeng
    Che, Wanxiang
    Sun, Zhongqian
    Liu, Ting
    Huang, Junzhou
    THIRTY-FOURTH AAAI CONFERENCE ON ARTIFICIAL INTELLIGENCE, THE THIRTY-SECOND INNOVATIVE APPLICATIONS OF ARTIFICIAL INTELLIGENCE CONFERENCE AND THE TENTH AAAI SYMPOSIUM ON EDUCATIONAL ADVANCES IN ARTIFICIAL INTELLIGENCE, 2020, 34 : 8838 - 8845
  • [15] RNN Transducers for Named Entity Recognition with constraints on alignment for understanding medical conversations
    Soltau, Hagen
    Shafran, Izhak
    Wang, Mingqiu
    El Shafey, Laurent
    INTERSPEECH 2022, 2022, : 1901 - 1905
  • [16] Parental literacy level and understanding of medical information
    Moon, RY
    Cheng, TL
    Patel, KM
    Baumhaft, K
    Scheidt, PC
    PEDIATRICS, 1998, 102 (02) : e25
  • [17] Information extraction from sound for medical telemonitoring
    Istrate, D
    Castelli, E
    Vacher, M
    Besacier, L
    Serignat, JF
    IEEE TRANSACTIONS ON INFORMATION TECHNOLOGY IN BIOMEDICINE, 2006, 10 (02): : 264 - 274
  • [18] Information extraction and summarization from medical documents
    Spyropoulos, CD
    Karkatetsis, V
    ARTIFICIAL INTELLIGENCE IN MEDICINE, 2005, 33 (02) : 107 - 110
  • [19] LLMs Accelerate Annotation for Medical Information Extraction
    Goel, Akshay
    Gueta, Almog
    Gilon, Omry
    Liu, Chang
    Erell, Sofia
    Lan Huong Nguyen
    Hao, Xiaohong
    Jaber, Bolous
    Reddy, Shashir
    Kartha, Rupesh
    Steiner, Jean
    Laish, Itay
    Feder, Amir
    MACHINE LEARNING FOR HEALTH, ML4H, VOL 225, 2023, 225 : 82 - 100
  • [20] An Extraction of Medical Information Based on Human Handwritings
    Bhaskoro, Susetyo Bagas
    Supangkat, Suhono Harso
    2014 INTERNATIONAL CONFERENCE ON INFORMATION TECHNOLOGY SYSTEMS AND INNOVATION (ICITSI), 2014, : 253 - 258