Qingxi Peng;Zhenjie Weng;Wei Wang;Xinyi Wang;Lan You
{"title":"A Collaborative Network-Based Retrieval Model for Open Source Domain Experts","authors":"Qingxi Peng;Zhenjie Weng;Wei Wang;Xinyi Wang;Lan You","doi":"10.1109/TBDATA.2024.3524829","DOIUrl":null,"url":null,"abstract":"Aiming at the problem that the GitHub platform only supports the retrieval of developers through their usernames and it is difficult to directly obtain developers' expertise information, this paper proposes an open source domain expert retrieval model (OSDERM) based on the network representation learning algorithm OSC2vec (Open Source Collaboration to Vector). The model mainly consists of two core parts: Expert Profiling and Expert Finding. Expert Profiling aims to enrich the expertise information in the search results by labeling the expertise of developers; while Expert Finding achieves rapid location of the most suitable domain experts through keyword matching, which greatly saves the time and effort of searching for experts in the open source community. Experiments using the GitHub ecological dataset show that the model outperforms existing comparative algorithms in discovering open source domain experts, and can provide an effective reference for enterprise recruitment","PeriodicalId":13106,"journal":{"name":"IEEE Transactions on Big Data","volume":"11 4","pages":"1720-1732"},"PeriodicalIF":5.7000,"publicationDate":"2025-01-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"0","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"IEEE Transactions on Big Data","FirstCategoryId":"94","ListUrlMain":"https://ieeexplore.ieee.org/document/10819979/","RegionNum":3,"RegionCategory":"计算机科学","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"Q1","JCRName":"COMPUTER SCIENCE, INFORMATION SYSTEMS","Score":null,"Total":0}
引用次数: 0
Abstract
Aiming at the problem that the GitHub platform only supports the retrieval of developers through their usernames and it is difficult to directly obtain developers' expertise information, this paper proposes an open source domain expert retrieval model (OSDERM) based on the network representation learning algorithm OSC2vec (Open Source Collaboration to Vector). The model mainly consists of two core parts: Expert Profiling and Expert Finding. Expert Profiling aims to enrich the expertise information in the search results by labeling the expertise of developers; while Expert Finding achieves rapid location of the most suitable domain experts through keyword matching, which greatly saves the time and effort of searching for experts in the open source community. Experiments using the GitHub ecological dataset show that the model outperforms existing comparative algorithms in discovering open source domain experts, and can provide an effective reference for enterprise recruitment
期刊介绍:
The IEEE Transactions on Big Data publishes peer-reviewed articles focusing on big data. These articles present innovative research ideas and application results across disciplines, including novel theories, algorithms, and applications. Research areas cover a wide range, such as big data analytics, visualization, curation, management, semantics, infrastructure, standards, performance analysis, intelligence extraction, scientific discovery, security, privacy, and legal issues specific to big data. The journal also prioritizes applications of big data in fields generating massive datasets.