{"title":"基于LDA的呼叫中心主题挖掘","authors":"Wenming Guo, Tianlang Deng","doi":"10.1109/ICNC.2014.6975947","DOIUrl":null,"url":null,"abstract":"Latent Dirichlet Allocation, which is a non-supervised learning method, can be used for topic detection, automatic text categorization, keyword extraction and so on. It only focuses on the text itself, not considering other external correlation properties. External association property refers to some structured attributes that correspondence with the text data, for example, a paper usually has several properties like authors, publishing time etc. A telephone call usually has several properties like caller number, call time etc. To iron out flaws; we propose an improved model A-LDA based LDA. We use data sets from telephone call centers (a kind of data centers in rapid growth) to experiment on topic detection. The topic results show that A-LDA with introduce of external correlation properties, compared with the traditional LDA, is decreased in perplexity value and has better generalization performance. At the same time, we can obtain the topic that external attributes contained.","PeriodicalId":208779,"journal":{"name":"2014 10th International Conference on Natural Computation (ICNC)","volume":"26 1","pages":"0"},"PeriodicalIF":0.0000,"publicationDate":"2014-12-08","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"2","resultStr":"{\"title\":\"Topic mining for call centers based on LDA\",\"authors\":\"Wenming Guo, Tianlang Deng\",\"doi\":\"10.1109/ICNC.2014.6975947\",\"DOIUrl\":null,\"url\":null,\"abstract\":\"Latent Dirichlet Allocation, which is a non-supervised learning method, can be used for topic detection, automatic text categorization, keyword extraction and so on. It only focuses on the text itself, not considering other external correlation properties. External association property refers to some structured attributes that correspondence with the text data, for example, a paper usually has several properties like authors, publishing time etc. A telephone call usually has several properties like caller number, call time etc. To iron out flaws; we propose an improved model A-LDA based LDA. We use data sets from telephone call centers (a kind of data centers in rapid growth) to experiment on topic detection. The topic results show that A-LDA with introduce of external correlation properties, compared with the traditional LDA, is decreased in perplexity value and has better generalization performance. At the same time, we can obtain the topic that external attributes contained.\",\"PeriodicalId\":208779,\"journal\":{\"name\":\"2014 10th International Conference on Natural Computation (ICNC)\",\"volume\":\"26 1\",\"pages\":\"0\"},\"PeriodicalIF\":0.0000,\"publicationDate\":\"2014-12-08\",\"publicationTypes\":\"Journal Article\",\"fieldsOfStudy\":null,\"isOpenAccess\":false,\"openAccessPdf\":\"\",\"citationCount\":\"2\",\"resultStr\":null,\"platform\":\"Semanticscholar\",\"paperid\":null,\"PeriodicalName\":\"2014 10th International Conference on Natural Computation (ICNC)\",\"FirstCategoryId\":\"1085\",\"ListUrlMain\":\"https://doi.org/10.1109/ICNC.2014.6975947\",\"RegionNum\":0,\"RegionCategory\":null,\"ArticlePicture\":[],\"TitleCN\":null,\"AbstractTextCN\":null,\"PMCID\":null,\"EPubDate\":\"\",\"PubModel\":\"\",\"JCR\":\"\",\"JCRName\":\"\",\"Score\":null,\"Total\":0}","platform":"Semanticscholar","paperid":null,"PeriodicalName":"2014 10th International Conference on Natural Computation (ICNC)","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.1109/ICNC.2014.6975947","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}
Latent Dirichlet Allocation, which is a non-supervised learning method, can be used for topic detection, automatic text categorization, keyword extraction and so on. It only focuses on the text itself, not considering other external correlation properties. External association property refers to some structured attributes that correspondence with the text data, for example, a paper usually has several properties like authors, publishing time etc. A telephone call usually has several properties like caller number, call time etc. To iron out flaws; we propose an improved model A-LDA based LDA. We use data sets from telephone call centers (a kind of data centers in rapid growth) to experiment on topic detection. The topic results show that A-LDA with introduce of external correlation properties, compared with the traditional LDA, is decreased in perplexity value and has better generalization performance. At the same time, we can obtain the topic that external attributes contained.