Identifying Textual Content Based on Thematic Analysis of Similar Texts in Big Data

V. Lytvyn, V. Vysotska, I. Peleshchak, T. Basyuk, V. Kovalchuk, Solomiya Kubinska, Bohdan Rusyn, L. Pohreliuk, L. Chyrun, T. Salo
{"title":"Identifying Textual Content Based on Thematic Analysis of Similar Texts in Big Data","authors":"V. Lytvyn, V. Vysotska, I. Peleshchak, T. Basyuk, V. Kovalchuk, Solomiya Kubinska, Bohdan Rusyn, L. Pohreliuk, L. Chyrun, T. Salo","doi":"10.1109/STC-CSIT.2019.8929808","DOIUrl":null,"url":null,"abstract":"The system designed for working with artistic texts in English is developed. It provides the opportunity to obtain the necessary information for the user, that is, the ordered data in the format Author - Title on the artwork, which correspond to the thematic query. This software tool operates to identify a plurality of content based on thematic analysis of texts. It is used as an application on a PC. Except for the main function the application does not perform additional actions. It consists of a part that serves as a user interface and a database that is located on the server. The input is an English-language text snippet that is similar to what the user wants to find. Output is a list of the names of works and the names of their authors, sorted by descending similarity.","PeriodicalId":271237,"journal":{"name":"2019 IEEE 14th International Conference on Computer Sciences and Information Technologies (CSIT)","volume":"9 1","pages":"0"},"PeriodicalIF":0.0000,"publicationDate":"2019-09-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"23","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"2019 IEEE 14th International Conference on Computer Sciences and Information Technologies (CSIT)","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.1109/STC-CSIT.2019.8929808","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}
引用次数: 23

Abstract

The system designed for working with artistic texts in English is developed. It provides the opportunity to obtain the necessary information for the user, that is, the ordered data in the format Author - Title on the artwork, which correspond to the thematic query. This software tool operates to identify a plurality of content based on thematic analysis of texts. It is used as an application on a PC. Except for the main function the application does not perform additional actions. It consists of a part that serves as a user interface and a database that is located on the server. The input is an English-language text snippet that is similar to what the user wants to find. Output is a list of the names of works and the names of their authors, sorted by descending similarity.
基于大数据相似文本主题分析的文本内容识别
开发了一个专门用于处理英语艺术文本的系统。它为用户提供了获取必要信息的机会,即与主题查询相对应的图稿上作者-标题格式的有序数据。该软件工具的操作是基于文本的主题分析来识别多个内容。在PC上作为应用程序使用。除了main函数外,应用程序不执行其他操作。它由作为用户界面的部分和位于服务器上的数据库组成。输入是与用户想要查找的内容相似的英语文本片段。输出是作品名称及其作者名称的列表,按相似性降序排序。
本文章由计算机程序翻译,如有差异,请以英文原文为准。
求助全文
约1分钟内获得全文 求助全文
来源期刊
自引率
0.00%
发文量
0
×
引用
GB/T 7714-2015
复制
MLA
复制
APA
复制
导出至
BibTeX EndNote RefMan NoteFirst NoteExpress
×
提示
您的信息不完整,为了账户安全,请先补充。
现在去补充
×
提示
您因"违规操作"
具体请查看互助需知
我知道了
×
提示
确定
请完成安全验证×
copy
已复制链接
快去分享给好友吧!
我知道了
右上角分享
点击右上角分享
0
联系我们:info@booksci.cn Book学术提供免费学术资源搜索服务,方便国内外学者检索中英文文献。致力于提供最便捷和优质的服务体验。 Copyright © 2023 布克学术 All rights reserved.
京ICP备2023020795号-1
ghs 京公网安备 11010802042870号
Book学术文献互助
Book学术文献互助群
群 号:604180095
Book学术官方微信