跳到主要导航 跳到搜索 跳到主要内容

Clinical large language model centered on electronic medical records

  • Yan Zhuang
  • , Bo Wang
  • , Chengliang Yin
  • , Junyan Zhang
  • , Fanqing Meng
  • , Jianfei Zhao
  • , Qingyong Su
  • , Xuan Zhao
  • , Xiuxing Li
  • , Ping Hu
  • , Shiyuan Liu
  • , Rilige Wu
  • , Yun Hua
  • , Wei Dong
  • , Bing Wei
  • , Li Zhang
  • , Lei Zheng
  • , João Conde*
  • , Ge Shi*
  • , Chong Feng*
  • Kunlun He*
*此作品的通讯作者
  • General Hospital of People's Liberation Army
  • Beijing Institute of Technology
  • The Second Affiliated Hospital of Xi’an Jiaotong University
  • NOVA University Lisbon

科研成果: 期刊稿件文章同行评审

摘要

In the quest to enhance medical consultation, our study introduces AI4Doctor, a sophisticated large-language model (LLM) tailored for the clinical domain. At the heart of AI4Doctor is an innovative integration strategy that synergizes distilled data extracted from electronic medical records (EMR) with empirical insights gathered from practicing physicians during the supervised fine-tuning. Although existing platforms offer informative responses, they fall short of replicating the nuanced decision-making processes of medical professionals, particularly in complex, integrative diagnostic scenarios. Motivated by the need to create a realistic medical practice environment, we propose that a combination of direct knowledge transfer from seasoned doctors and the strategic use of EMR can augment the abilities of LLM, enabling it to more closely mimic the clinical acumen of healthcare practitioners. To navigate the complexities of merging diverse instructional sources, we employ a curriculum learning approach during the fine-tuning process. Moreover, we advance our model’s performance by developing a reward system that incentivizes the alignment of the LLM’s outputs with the valuable attributes inherent in both doctors’ expertise, including diagnostic priors, risk thresholds, and heuristic saliencies accumulated from practice and EMR data. This is achieved through a novel reinforcement-learning approach. Besides, we introduce a new benchmark involving a comparative evaluation. We utilize a subjective evaluation system wherein experts critically assess the responses from a professional perspective as well. Our research underscores the potential of this hybrid model to serve as a robust tool in medical consultations, bridging the gap between artificial intelligence and real-world clinical practice. (Figure presented.)

源语言英语
期刊论文编号511
期刊npj Digital Medicine
9
1
DOI
出版状态已出版 - 12月 2026
已对外发布

学术指纹

探究 'Clinical large language model centered on electronic medical records' 的科研主题。它们共同构成独一无二的学术指纹。

引用此