Content moderation assistance through image caption generation

IF 4.3

Intelligent Systems with Applications Pub Date : 2025-02-16 DOI:10.1016/j.iswa.2025.200489

Liam Kearns

{"title":"Content moderation assistance through image caption generation","authors":"Liam Kearns","doi":"10.1016/j.iswa.2025.200489","DOIUrl":null,"url":null,"abstract":"<div><div>The rapid growth in digital media creation has led to an increased challenge in content moderation. Manual and automated moderation are susceptible to risks associated with a slower response time and false positives arising from unpredictable user inputs respectively. Image caption generation has been suggested as a viable content moderation tool, but there is a lack of real world deployment in this context. In this work, a collaborative approach is taken, where a machine learning model is used to assist human moderators in the approval and rejection of media within a scavenger hunt game. The proposed model is trained on the Flickr30k and MS Coco datasets to generate captions for images. The results demonstrate a 13% reduction in review times, indicating that human–machine collaboration contributes to mitigating the risk of unsustainable review backlog growth. Furthermore, fine-tuning the model led to a 28% reduction in review times when compared to the untuned model. Notably, this paper contributes to knowledge by demonstrating caption generation as a viable content moderation tool in addition to its sensitivity to accurate captions, whereby false positives risk a deterioration in moderator response time.</div></div>","PeriodicalId":100684,"journal":{"name":"Intelligent Systems with Applications","volume":"25 ","pages":"Article 200489"},"PeriodicalIF":4.3000,"publicationDate":"2025-02-16","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"0","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"Intelligent Systems with Applications","FirstCategoryId":"1085","ListUrlMain":"https://www.sciencedirect.com/science/article/pii/S2667305325000158","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}

引用次数: 0

Abstract

The rapid growth in digital media creation has led to an increased challenge in content moderation. Manual and automated moderation are susceptible to risks associated with a slower response time and false positives arising from unpredictable user inputs respectively. Image caption generation has been suggested as a viable content moderation tool, but there is a lack of real world deployment in this context. In this work, a collaborative approach is taken, where a machine learning model is used to assist human moderators in the approval and rejection of media within a scavenger hunt game. The proposed model is trained on the Flickr30k and MS Coco datasets to generate captions for images. The results demonstrate a 13% reduction in review times, indicating that human–machine collaboration contributes to mitigating the risk of unsustainable review backlog growth. Furthermore, fine-tuning the model led to a 28% reduction in review times when compared to the untuned model. Notably, this paper contributes to knowledge by demonstrating caption generation as a viable content moderation tool in addition to its sensitivity to accurate captions, whereby false positives risk a deterioration in moderator response time.

Abstract Image

查看原文本刊更多论文

通过图像标题生成帮助内容审核

数字媒体创作的快速增长导致内容审核面临越来越大的挑战。手动和自动审核分别容易受到响应时间较慢和由不可预测的用户输入引起的误报相关风险的影响。图像标题生成已被建议作为一种可行的内容审核工具，但在此上下文中缺乏实际部署。在这项工作中，采用了一种协作方法，其中使用机器学习模型来协助人类版主批准和拒绝寻宝游戏中的媒体。该模型在Flickr30k和MS Coco数据集上进行训练，生成图像的说明文字。结果显示审查时间减少了13%，表明人机协作有助于减轻不可持续的审查积压增长的风险。此外，与未调整的模型相比，对模型进行微调可以减少28%的审查时间。值得注意的是，本文通过证明标题生成是一种可行的内容审核工具，以及它对准确标题的敏感性，从而为知识做出贡献，因此假阳性可能会导致版主响应时间的恶化。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

求助全文

约1分钟内获得全文求助全文

来源期刊

Intelligent Systems with Applications

CiteScore

5.60

自引率

0.00%

发文量