预印本 · 建模 / 计算研究
使用大语言模型从心导管报告中自动识别复杂经皮冠状动脉介入治疗
Automated Identification of Complex Percutaneous Coronary Intervention from Cardiac Catheterization Reports using Large Language Models
作者:Nilay Bhatt, Fred Warner, Jennifer Miao, Ravi Thakker, Golsa Joodi, Pablo Cantero-Schaffer, Bobak Mortazavi, Chenxi Huang, Harlan M Krumholz, Karthik Murugiah
medRxiv · 2026年9月16日 · Bhatt 等 10 位作者
不需要生物学背景,多打比方
正在获取全文并生成讲解(拿不到全文就依据摘要)…
已等待 0 秒大约需要 10–20 秒
可以先看别的,做好了会自动出现在这里。
这篇还没有动画
动画会把研究的流程、作用机制和关键结果一步一步演示出来,每一步都标明出自原文哪里。制作大约需要一两分钟。
摘要Abstract
INTRODUCTION: Manual abstraction of complex percutaneous coronary intervention (PCI) variables from cardiac catheterization reports is labor-intensive and limits scalable cardiovascular research. We evaluated open-weight large language models (LLMs) for automated complex PCI phenotyping.
METHODS: We evaluated three LLMs (Llama 3.3 70B, Meditron-7B, and BioMistral-7B) using manually annotated catheterization reports from three hospitals within Yale New Haven Health. Models identified PCI reports and extracted six complex PCI features: 3 vessels treated, ≥3 lesions treated, bifurcation PCI with two stents, chronic total occlusion, ≥3 stents, and total stent length ≥60 mm.
RESULTS: Among 1,412 clinical notes, 596 were PCI reports. Llama 3.3 70B outperformed the smaller domain-specific models across most tasks. For PCI identification, Llama 3.3 70B achieved 100.0% sensitivity, 93.8% specificity, 96.4% accuracy, and 95.9% F1 score. Among 590 evaluable PCI reports (excluding 6 indeterminable cases due to missing variables) for complex PCI classification, Llama 3.3 70B achieved 97.7% sensitivity, 80.1% specificity, 57.6% positive predictive value, 99.2% negative predictive value, 83.9% accuracy, and 72.5% F1 score. Performance was higher for explicitly documented variables, including stent number and length, and lower for variables requiring interpretation across procedural details, including lesion count, bifurcation PCI, and chronic total occlusion PCI. Llama 3.3 70B had the highest accuracy at each site for complex PCI classification but significant site-level heterogeneity was observed.
CONCLUSION: A high-capacity open-weight LLM accurately extracted complex PCI variables from unstructured reports and outperformed smaller domain-specific models. These findings support the potential use of locally deployable LLMs for scalable automated PCI phenotyping.
还没有查过关联研究
我会去找这篇研究之前的基础工作、做类似事情的研究,以及之后引用它的研究,并说明每篇为什么相关。