Enhancing Biomedical Text Summarization and Question-Answering: On the Utility of Domain-Specific Pre-Training

Conference and Labs of the Evaluation Forum Pub Date : 2023-07-10 DOI:10.48550/arXiv.2307.04412

Dima Galat, Marian-Andrei Rizoiu

引用次数: 1

Abstract

Biomedical summarization requires large datasets to train for text generation. We show that while transfer learning offers a viable option for addressing this challenge, an in-domain pre-training does not always offer advantages in a BioASQ summarization task. We identify a suitable model architecture and use it to show a benefit of a general-domain pre-training followed by a task-specific fine-tuning in the context of a BioASQ summarization task, leading to a novel three-step fine-tuning approach that works with only a thousand in-domain examples. Our results indicate that a Large Language Model without domain-specific pre-training can have a significant edge in some domain-specific biomedical text generation tasks.

查看原文本刊更多论文

增强生物医学文本摘要与问答:基于特定领域预训练的应用

生物医学摘要需要大量的数据集来训练文本生成。我们表明，虽然迁移学习为解决这一挑战提供了一个可行的选择，但域内预训练并不总是在BioASQ总结任务中提供优势。我们确定了一个合适的模型架构，并使用它来显示一般领域预训练的好处，然后在BioASQ总结任务的上下文中进行特定于任务的微调，从而产生了一种新的三步微调方法，该方法仅适用于一千个域内示例。我们的研究结果表明，没有特定领域预训练的大型语言模型在某些特定领域的生物医学文本生成任务中具有显著的优势。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

求助全文

约1分钟内获得全文求助全文

来源期刊

Conference and Labs of the Evaluation Forum

自引率

0.00%

发文量