医学伦理研究助手
前沿
论文精读

预印本 · 建模 / 计算研究

使用大语言模型从心导管报告中自动识别复杂经皮冠状动脉介入治疗

medRxiv · 2026年9月16日 · Bhatt 等 10 位作者

这是预印本,还没有经过同行评议预印本是作者先公开的稿件,结论可能在正式发表前被修改。
问这篇
一分钟了解要点开源大模型Llama 3.3 70B从1412份导管报告中识别PCI及复杂特征,准确率最高达96.4%。结果Llama 3.3 70B在PCI识别上达100%敏感度、93.8%特异度、96.4%准确度、95.9% F1值;复杂PCI分类在590份可评估报告中敏感度97.7%、特异度80.1%、准确度83.9%、F1值72.5%,对明确记录的变量(如支架数量和长度)表现更好,对需要跨细节解读的变量(如病变数、分叉PCI、慢性完全闭塞)较差,且各中心间存在异质性。

不需要生物学背景,多打比方

正在获取全文并生成讲解(拿不到全文就依据摘要)…

已等待 0 秒大约需要 10–20 秒

可以先看别的,做好了会自动出现在这里。

这篇还没有动画

动画会把研究的流程、作用机制和关键结果一步一步演示出来,每一步都标明出自原文哪里。制作大约需要一两分钟。

目前只拿到了摘要全文暂时拿不到(可能不是免费全文)。下面是论文摘要。

摘要Abstract

摘要第 1 段问这一段

INTRODUCTION: Manual abstraction of complex percutaneous coronary intervention (PCI) variables from cardiac catheterization reports is labor-intensive and limits scalable cardiovascular research. We evaluated open-weight large language models (LLMs) for automated complex PCI phenotyping.

摘要第 2 段问这一段

METHODS: We evaluated three LLMs (Llama 3.3 70B, Meditron-7B, and BioMistral-7B) using manually annotated catheterization reports from three hospitals within Yale New Haven Health. Models identified PCI reports and extracted six complex PCI features: 3 vessels treated, ≥3 lesions treated, bifurcation PCI with two stents, chronic total occlusion, ≥3 stents, and total stent length ≥60 mm.

摘要第 3 段问这一段

RESULTS: Among 1,412 clinical notes, 596 were PCI reports. Llama 3.3 70B outperformed the smaller domain-specific models across most tasks. For PCI identification, Llama 3.3 70B achieved 100.0% sensitivity, 93.8% specificity, 96.4% accuracy, and 95.9% F1 score. Among 590 evaluable PCI reports (excluding 6 indeterminable cases due to missing variables) for complex PCI classification, Llama 3.3 70B achieved 97.7% sensitivity, 80.1% specificity, 57.6% positive predictive value, 99.2% negative predictive value, 83.9% accuracy, and 72.5% F1 score. Performance was higher for explicitly documented variables, including stent number and length, and lower for variables requiring interpretation across procedural details, including lesion count, bifurcation PCI, and chronic total occlusion PCI. Llama 3.3 70B had the highest accuracy at each site for complex PCI classification but significant site-level heterogeneity was observed.

摘要第 4 段问这一段

CONCLUSION: A high-capacity open-weight LLM accurately extracted complex PCI variables from unstructured reports and outperformed smaller domain-specific models. These findings support the potential use of locally deployable LLMs for scalable automated PCI phenotyping.

这篇对您:
讲解或动画有问题: