Understanding Medical Conversations: Rich Transcription, Confidence Scores & Information Extraction

被引:0
|
作者
Soltau, Hagen [1 ]
Wang, Mingqiu [1 ]
Shafran, Izhak [1 ]
El Shafey, Laurent [1 ]
机构
[1] Google, Mountain View, CA 94043 USA
来源
INTERSPEECH 2021 | 2021年
关键词
SPEECH RECOGNITION;
D O I
暂无
中图分类号
R36 [病理学]; R76 [耳鼻咽喉科学];
学科分类号
100104 ; 100213 ;
摘要
In this paper, we describe novel components for extracting clinically relevant information from medical conversations which will be available as Google APIs. We describe a transformerbased Recurrent Neural Network Transducer (RNN-T) model tailored for long-form audio, which can produce rich transcriptions including speaker segmentation, speaker role labeling, punctuation and capitalization. On a representative test set, we compare performance of RNN-T models with different encoders, units and streaming constraints. Our transformer-based streaming model performs at about 20% WER on the ASR task, 6% WDER on the diarization task, 43% SER on periods, 52% SER on commas, 43% SER on question marks and 30% SER on capitalization. Our recognizer is paired with a confidence model that utilizes both acoustic and lexical features from the recognizer. The model performs at about 0.37 NCE. Finally, we describe a RNN-T based tagging model. The performance of the model depends on the ontologies, with F-scores of 0.90 for medications, 0.76 for symptoms, 0.75 for conditions, 0.76 for diagnosis, and 0.61 for treatments. While there is still room for improvement, our results suggest that these models are sufficiently accurate for practical applications.
引用
收藏
页码:4418 / 4422
页数:5
相关论文
共 50 条
  • [1] Failure Prediction in 2D Document Information Extraction with Calibrated Confidence Scores
    Kivimaki, Juhani
    Lebedev, Aleksey
    Nurminen, Jukka K.
    2023 IEEE 47TH ANNUAL COMPUTERS, SOFTWARE, AND APPLICATIONS CONFERENCE, COMPSAC, 2023, : 193 - 202
  • [2] Towards Understanding ASR Error Correction for Medical Conversations
    Mani, Anirudh
    Palaskar, Shruti
    Konam, Sandeep
    NATURAL LANGUAGE PROCESSING FOR MEDICAL CONVERSATIONS, 2020, : 7 - 11
  • [3] Information extraction and webpage understanding
    Sharmila Begum, M.
    Dinesh, L.
    Aruna, P.
    International Journal of Computer Science Issues, 2011, 8 (6 6-3): : 304 - 308
  • [4] Enhancing Medical Learners' Knowledge of, Comfort and Confidence in Holding Serious Illness Conversations
    Tam, Vivian
    You, John J.
    Bernacki, Rachelle
    AMERICAN JOURNAL OF HOSPICE & PALLIATIVE MEDICINE, 2019, 36 (12): : 1096 - 1104
  • [5] Information Extraction in the Medical Domain
    Ghoulam, Aicha
    Barigou, Fatiha
    Belalem, Ghalem
    JOURNAL OF INFORMATION TECHNOLOGY RESEARCH, 2015, 8 (02) : 1 - 15
  • [6] CRMI: Confidence-Rich Mutual Information for Information-Theoretic Mapping
    Xu, Yang
    Zheng, Ronghao
    Liu, Meiqin
    Zhang, Senlin
    IEEE ROBOTICS AND AUTOMATION LETTERS, 2021, 6 (04) : 6434 - 6441
  • [7] Semantic information extraction in medical information systems
    Holzinger, Andreas
    Geierhofer, Regina
    Errath, Maximilian
    Informatik-Spektrum, 2007, 30 (02) : 69 - 78
  • [8] Causal Domain Adaptation for Information Extraction from Complex Conversations
    Li, Xue
    SEMANTIC WEB: ESWC 2022 SATELLITE EVENTS, 2022, 13384 : 189 - 198
  • [9] Japanese broadcast news transcription and information extraction
    Furui, S
    Ohtsuki, K
    Zhang, ZP
    COMMUNICATIONS OF THE ACM, 2000, 43 (02) : 71 - 73
  • [10] Multi-aspect Understanding with Cooperative Graph Attention Networks for Medical Dialogue Information Extraction
    Lin, Rui
    Fan, Jing
    Wu, Haifeng
    ACM TRANSACTIONS ON INTELLIGENT SYSTEMS AND TECHNOLOGY, 2023, 14 (06)