Creating and analysing a multimodal corpus of news texts with Google Cloud Vision's automatic image tagger

Applied Corpus Linguistics Pub Date : 2023-04-01 DOI:10.1016/j.acorp.2023.100043

Paul Baker, Luke Collins

引用次数: 0

Abstract

This study describes the creation and analysis of a small multimodal corpus of British news articles about obesity, where tags were assigned to images in the articles using the automatic tagger Google Cloud Vision. In order to illustrate the potential for analysis of image tags, the corpus analysis tool WordSmith was used to identify differences between newspapers in the ways that obesity was framed. Three forms of analysis were carried out – the first simply compared keywords across the newspapers, the second examined key visual tags and their collocates associated with each newspaper, while the third incorporated a combined analysis of words and image tags. The three analyses produced complementary findings, indicating the value in using Google Cloud Vision in creating and analysing multimodal corpora. The paper ends by reflecting on the method undertaken, while considering how additional research could improve our understanding of image tagging.

查看原文本刊更多论文

使用谷歌云视觉的自动图像标注器创建和分析新闻文本的多模态语料库

本研究描述了一个关于肥胖的英国新闻文章的小型多模态语料库的创建和分析，其中使用自动标记器Google Cloud Vision为文章中的图像分配标签。为了说明图像标签分析的潜力，语料库分析工具WordSmith被用来识别报纸在肥胖框架方面的差异。研究人员进行了三种形式的分析——第一种简单地比较了报纸上的关键词，第二种检查了关键的视觉标签及其与每份报纸相关的搭配，而第三种结合了文字和图像标签的综合分析。这三个分析产生了互补的结果，表明了使用谷歌云视觉在创建和分析多模态语料库中的价值。本文最后反思了所采用的方法，同时考虑了如何进一步研究可以提高我们对图像标记的理解。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

求助全文

约1分钟内获得全文求助全文

来源期刊