基于卷积神经网络的马来语脏话分类

2021 IEEE International Conference on Signal and Image Processing Applications (ICSIPA) Pub Date : 2021-09-13 DOI:10.1109/ICSIPA52582.2021.9576781

A. Wazir, H. A. Karim, Nouar Aldahoul, M. F. A. Fauzi, Sarina Mansor, Mohd Haris Lye Abdullah, Hor Sui Lyn, Tabibah Zainab Zulkifli

{"title":"基于卷积神经网络的马来语脏话分类","authors":"A. Wazir, H. A. Karim, Nouar Aldahoul, M. F. A. Fauzi, Sarina Mansor, Mohd Haris Lye Abdullah, Hor Sui Lyn, Tabibah Zainab Zulkifli","doi":"10.1109/ICSIPA52582.2021.9576781","DOIUrl":null,"url":null,"abstract":"Foul language exists in films, video-sharing platforms, and social media platforms, which increase the risk of a viewer to be exposed to large number of profane words that have negative personal and social impact. This work proposes a CNN-based spoken Malay foul words recognition to establish the base of spoken foul terms detection for monitoring and censorship purpose. A novel foul speech containing 1512 samples are collected, processed, and annotated. The dataset then has been converted into spectral representation of Mel-spectrogram images to be used as an input to CNN model. This research proposes a lightweight CNN model with only six convolutional layers and small size filters to minimize the computational cost. The proposed model’s performance affirms the viability of the proposed visual-based classification method using CNN by achieving an average Malay foul speech terms classification accuracy of 86.50%, precision of 88.68%, and F-score of 86.83. The class of normal conversational class outperformed the class of foul words due to data imbalance and rarity of foul speech samples compared to normal speech terms.","PeriodicalId":326688,"journal":{"name":"2021 IEEE International Conference on Signal and Image Processing Applications (ICSIPA)","volume":"1 1","pages":"0"},"PeriodicalIF":0.0000,"publicationDate":"2021-09-13","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"1","resultStr":"{\"title\":\"Spoken Malay Profanity Classification Using Convolutional Neural Network\",\"authors\":\"A. Wazir, H. A. Karim, Nouar Aldahoul, M. F. A. Fauzi, Sarina Mansor, Mohd Haris Lye Abdullah, Hor Sui Lyn, Tabibah Zainab Zulkifli\",\"doi\":\"10.1109/ICSIPA52582.2021.9576781\",\"DOIUrl\":null,\"url\":null,\"abstract\":\"Foul language exists in films, video-sharing platforms, and social media platforms, which increase the risk of a viewer to be exposed to large number of profane words that have negative personal and social impact. This work proposes a CNN-based spoken Malay foul words recognition to establish the base of spoken foul terms detection for monitoring and censorship purpose. A novel foul speech containing 1512 samples are collected, processed, and annotated. The dataset then has been converted into spectral representation of Mel-spectrogram images to be used as an input to CNN model. This research proposes a lightweight CNN model with only six convolutional layers and small size filters to minimize the computational cost. The proposed model’s performance affirms the viability of the proposed visual-based classification method using CNN by achieving an average Malay foul speech terms classification accuracy of 86.50%, precision of 88.68%, and F-score of 86.83. The class of normal conversational class outperformed the class of foul words due to data imbalance and rarity of foul speech samples compared to normal speech terms.\",\"PeriodicalId\":326688,\"journal\":{\"name\":\"2021 IEEE International Conference on Signal and Image Processing Applications (ICSIPA)\",\"volume\":\"1 1\",\"pages\":\"0\"},\"PeriodicalIF\":0.0000,\"publicationDate\":\"2021-09-13\",\"publicationTypes\":\"Journal Article\",\"fieldsOfStudy\":null,\"isOpenAccess\":false,\"openAccessPdf\":\"\",\"citationCount\":\"1\",\"resultStr\":null,\"platform\":\"Semanticscholar\",\"paperid\":null,\"PeriodicalName\":\"2021 IEEE International Conference on Signal and Image Processing Applications (ICSIPA)\",\"FirstCategoryId\":\"1085\",\"ListUrlMain\":\"https://doi.org/10.1109/ICSIPA52582.2021.9576781\",\"RegionNum\":0,\"RegionCategory\":null,\"ArticlePicture\":[],\"TitleCN\":null,\"AbstractTextCN\":null,\"PMCID\":null,\"EPubDate\":\"\",\"PubModel\":\"\",\"JCR\":\"\",\"JCRName\":\"\",\"Score\":null,\"Total\":0}","platform":"Semanticscholar","paperid":null,"PeriodicalName":"2021 IEEE International Conference on Signal and Image Processing Applications (ICSIPA)","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.1109/ICSIPA52582.2021.9576781","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}

引用次数: 1

摘要

脏话存在于电影、视频分享平台和社交媒体平台中，这增加了观众接触大量脏话的风险，这些脏话对个人和社会都有负面影响。本研究提出了一种基于cnn的马来语口语脏话识别方法，以建立口语脏话检测的基础，用于监控和审查目的。收集、处理并注释了一种包含1512个样本的新颖污言秽语。然后将数据集转换为mel光谱图图像的光谱表示，用作CNN模型的输入。本研究提出了一种轻量级的CNN模型，只有6个卷积层和小尺寸滤波器，以最小化计算成本。该模型的性能证实了本文提出的基于CNN的基于视觉的分类方法的可行性，马来语脏话术语的平均分类准确率为86.50%，精度为88.68%，f分为86.83。由于数据不平衡和脏话样本的稀有性，正常会话类的表现优于脏话类。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

查看原文本刊更多论文

Spoken Malay Profanity Classification Using Convolutional Neural Network

Foul language exists in films, video-sharing platforms, and social media platforms, which increase the risk of a viewer to be exposed to large number of profane words that have negative personal and social impact. This work proposes a CNN-based spoken Malay foul words recognition to establish the base of spoken foul terms detection for monitoring and censorship purpose. A novel foul speech containing 1512 samples are collected, processed, and annotated. The dataset then has been converted into spectral representation of Mel-spectrogram images to be used as an input to CNN model. This research proposes a lightweight CNN model with only six convolutional layers and small size filters to minimize the computational cost. The proposed model’s performance affirms the viability of the proposed visual-based classification method using CNN by achieving an average Malay foul speech terms classification accuracy of 86.50%, precision of 88.68%, and F-score of 86.83. The class of normal conversational class outperformed the class of foul words due to data imbalance and rarity of foul speech samples compared to normal speech terms.

求助全文

通过发布文献求助，成功后即可免费获取论文全文。去求助

来源期刊

2021 IEEE International Conference on Signal and Image Processing Applications (ICSIPA)

自引率

0.00%

发文量