Hyperbolic graph attention network fusing long-context for technical keyphrase extraction

IF 14.7 1区计算机科学 Q1 COMPUTER SCIENCE, ARTIFICIAL INTELLIGENCE

Information Fusion Pub Date : 2025-03-07 DOI:10.1016/j.inffus.2025.103061

Yushan Zhao , Kuan-Ching Li , Shunxiang Zhang , Tongzhou Ye

{"title":"Hyperbolic graph attention network fusing long-context for technical keyphrase extraction","authors":"Yushan Zhao , Kuan-Ching Li , Shunxiang Zhang , Tongzhou Ye","doi":"10.1016/j.inffus.2025.103061","DOIUrl":null,"url":null,"abstract":"<div><div>Technical Keyphrase Extraction (TKE) is crucial for summarizing the core content of scientific and technical texts. Existing keyphrase extraction models typically focus on calculating phrase and sentence correlations that can limit their ability to understand long contexts and uncover hierarchical semantic information, leading to biased results. To address these limitations, a hyperbolic graph technical attention network is designed and applied to a novel unsupervised Technical KeyPhrase Extraction (TKPE) model, achieving the fusion of complex hierarchical semantic representations and long-context information by constructing global embeddings of the technical text in hyperbolic space for high-fidelity representation with minimal dimensions. A technical attention score is calculated based on technical terminology degree and hierarchical relevance to guide the extraction process. Additionally, the network utilizes geodesic variations between embedded nodes to reveal meaningful hierarchical clustering relationships, thus enabling semantic structural understanding of technical text data and efficient extraction of the most relevant technical keyphrases. This work exploits the long-context understanding capability of large language models to generate candidate phrases guided by an effective prompt template that reduces information loss when importing candidate phrases in a hyperbolic graph attention network. Experiments performed on benchmark technical datasets demonstrate that the proposed model outperforms recent state-of-the-art baseline keyphrase extraction models.</div></div>","PeriodicalId":50367,"journal":{"name":"Information Fusion","volume":"120 ","pages":"Article 103061"},"PeriodicalIF":14.7000,"publicationDate":"2025-03-07","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"0","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"Information Fusion","FirstCategoryId":"94","ListUrlMain":"https://www.sciencedirect.com/science/article/pii/S1566253525001344","RegionNum":1,"RegionCategory":"计算机科学","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"Q1","JCRName":"COMPUTER SCIENCE, ARTIFICIAL INTELLIGENCE","Score":null,"Total":0}

引用次数: 0

Abstract

Technical Keyphrase Extraction (TKE) is crucial for summarizing the core content of scientific and technical texts. Existing keyphrase extraction models typically focus on calculating phrase and sentence correlations that can limit their ability to understand long contexts and uncover hierarchical semantic information, leading to biased results. To address these limitations, a hyperbolic graph technical attention network is designed and applied to a novel unsupervised Technical KeyPhrase Extraction (TKPE) model, achieving the fusion of complex hierarchical semantic representations and long-context information by constructing global embeddings of the technical text in hyperbolic space for high-fidelity representation with minimal dimensions. A technical attention score is calculated based on technical terminology degree and hierarchical relevance to guide the extraction process. Additionally, the network utilizes geodesic variations between embedded nodes to reveal meaningful hierarchical clustering relationships, thus enabling semantic structural understanding of technical text data and efficient extraction of the most relevant technical keyphrases. This work exploits the long-context understanding capability of large language models to generate candidate phrases guided by an effective prompt template that reduces information loss when importing candidate phrases in a hyperbolic graph attention network. Experiments performed on benchmark technical datasets demonstrate that the proposed model outperforms recent state-of-the-art baseline keyphrase extraction models.

查看原文本刊更多论文

求助全文

约1分钟内获得全文求助全文

来源期刊

Information Fusion 工程技术-计算机：理论方法

CiteScore

33.20

自引率

4.30%

发文量

161

审稿时长

7.9 months

期刊介绍： Information Fusion serves as a central platform for showcasing advancements in multi-sensor, multi-source, multi-process information fusion, fostering collaboration among diverse disciplines driving its progress. It is the leading outlet for sharing research and development in this field, focusing on architectures, algorithms, and applications. Papers dealing with fundamental theoretical analyses as well as those demonstrating their application to real-world problems will be welcome.