Esla Timothy Anzaku, Hyesoo Hong, Jin-Woo Park, Wonjun Yang, Kangmin Kim, Jongbum Won, Deshika Vinoshani Kumari Herath, Arnout Van Messem, W. D. Neve
{"title":"Leveraging Human-Machine Interactions for Computer Vision Dataset Quality Enhancement","authors":"Esla Timothy Anzaku, Hyesoo Hong, Jin-Woo Park, Wonjun Yang, Kangmin Kim, Jongbum Won, Deshika Vinoshani Kumari Herath, Arnout Van Messem, W. D. Neve","doi":"10.48550/arXiv.2401.17736","DOIUrl":"https://doi.org/10.48550/arXiv.2401.17736","url":null,"abstract":"Large-scale datasets for single-label multi-class classification, such as emph{ImageNet-1k}, have been instrumental in advancing deep learning and computer vision. However, a critical and often understudied aspect is the comprehensive quality assessment of these datasets, especially regarding potential multi-label annotation errors. In this paper, we introduce a lightweight, user-friendly, and scalable framework that synergizes human and machine intelligence for efficient dataset validation and quality enhancement. We term this novel framework emph{Multilabelfy}. Central to Multilabelfy is an adaptable web-based platform that systematically guides annotators through the re-evaluation process, effectively leveraging human-machine interactions to enhance dataset quality. By using Multilabelfy on the ImageNetV2 dataset, we found that approximately $47.88%$ of the images contained at least two labels, underscoring the need for more rigorous assessments of such influential datasets. Furthermore, our analysis showed a negative correlation between the number of potential labels per image and model top-1 accuracy, illuminating a crucial factor in model evaluation and selection. Our open-source framework, Multilabelfy, offers a convenient, lightweight solution for dataset enhancement, emphasizing multi-label proportions. This study tackles major challenges in dataset integrity and provides key insights into model performance evaluation. Moreover, it underscores the advantages of integrating human expertise with machine capabilities to produce more robust models and trustworthy data development. The source code for Multilabelfy will be available at https://github.com/esla/Multilabelfy. keywords{Computer Vision and Dataset Quality Enhancement and Dataset Validation and Human-Computer Interaction and Multi-label Annotation.}","PeriodicalId":224881,"journal":{"name":"International Conference on Intelligent Human Computer Interaction","volume":"812 ","pages":"295-309"},"PeriodicalIF":0.0,"publicationDate":"2024-01-31","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":null,"resultStr":null,"platform":"Semanticscholar","paperid":"140479353","PeriodicalName":null,"FirstCategoryId":null,"ListUrlMain":null,"RegionNum":0,"RegionCategory":"","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":"","EPubDate":null,"PubModel":null,"JCR":null,"JCRName":null,"Score":null,"Total":0}
Arvind W. Kiwelekar, Swanand Navandar, Dharmendra Yadav
{"title":"A Two-Systems Perspective for Computational Thinking","authors":"Arvind W. Kiwelekar, Swanand Navandar, Dharmendra Yadav","doi":"10.1007/978-3-030-68449-5_1","DOIUrl":"https://doi.org/10.1007/978-3-030-68449-5_1","url":null,"abstract":"","PeriodicalId":224881,"journal":{"name":"International Conference on Intelligent Human Computer Interaction","volume":"2 1","pages":"0"},"PeriodicalIF":0.0,"publicationDate":"2020-12-06","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":null,"resultStr":null,"platform":"Semanticscholar","paperid":"123865961","PeriodicalName":null,"FirstCategoryId":null,"ListUrlMain":null,"RegionNum":0,"RegionCategory":"","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":"","EPubDate":null,"PubModel":null,"JCR":null,"JCRName":null,"Score":null,"Total":0}
{"title":"The Commodity Ecology Mobile (CEM) Platform Illustrates Ten Design Points for Achieving a Deep Deliberation in Sustainable Development Goal #12","authors":"M. Whitaker","doi":"10.1007/978-3-030-68449-5_41","DOIUrl":"https://doi.org/10.1007/978-3-030-68449-5_41","url":null,"abstract":"","PeriodicalId":224881,"journal":{"name":"International Conference on Intelligent Human Computer Interaction","volume":"29 1","pages":"0"},"PeriodicalIF":0.0,"publicationDate":"2020-11-24","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":null,"resultStr":null,"platform":"Semanticscholar","paperid":"123083436","PeriodicalName":null,"FirstCategoryId":null,"ListUrlMain":null,"RegionNum":0,"RegionCategory":"","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":"","EPubDate":null,"PubModel":null,"JCR":null,"JCRName":null,"Score":null,"Total":0}
Na Yeon Han, S. W. Seong, Jihye Ryu, Hyeonsang Hwang, Jinoo Joung, J. Lee, E. Lee
{"title":"Authentication of Facial Images with Masks Using Periocular Biometrics","authors":"Na Yeon Han, S. W. Seong, Jihye Ryu, Hyeonsang Hwang, Jinoo Joung, J. Lee, E. Lee","doi":"10.1007/978-3-030-68452-5_34","DOIUrl":"https://doi.org/10.1007/978-3-030-68452-5_34","url":null,"abstract":"","PeriodicalId":224881,"journal":{"name":"International Conference on Intelligent Human Computer Interaction","volume":"234 1","pages":"0"},"PeriodicalIF":0.0,"publicationDate":"2020-11-24","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":null,"resultStr":null,"platform":"Semanticscholar","paperid":"132695700","PeriodicalName":null,"FirstCategoryId":null,"ListUrlMain":null,"RegionNum":0,"RegionCategory":"","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":"","EPubDate":null,"PubModel":null,"JCR":null,"JCRName":null,"Score":null,"Total":0}
{"title":"A Method for Localizing and Grasping Objects in a Picking Robot System Using Kinect Camera","authors":"T. Nguyen, Trung Trong Nguyen, T. Tran","doi":"10.1007/978-3-030-68452-5_2","DOIUrl":"https://doi.org/10.1007/978-3-030-68452-5_2","url":null,"abstract":"","PeriodicalId":224881,"journal":{"name":"International Conference on Intelligent Human Computer Interaction","volume":"76 1","pages":"0"},"PeriodicalIF":0.0,"publicationDate":"2020-11-24","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":null,"resultStr":null,"platform":"Semanticscholar","paperid":"117230733","PeriodicalName":null,"FirstCategoryId":null,"ListUrlMain":null,"RegionNum":0,"RegionCategory":"","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":"","EPubDate":null,"PubModel":null,"JCR":null,"JCRName":null,"Score":null,"Total":0}
Akash K Rao, Jibraan Singh Chahal, Sushil Chandra, V. Dutt
{"title":"Virtual-Reality Training Under Varying Degrees of Task Difficulty in a Complex Search-and-Shoot Scenario","authors":"Akash K Rao, Jibraan Singh Chahal, Sushil Chandra, V. Dutt","doi":"10.1007/978-3-030-44689-5_22","DOIUrl":"https://doi.org/10.1007/978-3-030-44689-5_22","url":null,"abstract":"","PeriodicalId":224881,"journal":{"name":"International Conference on Intelligent Human Computer Interaction","volume":"2 1","pages":"0"},"PeriodicalIF":0.0,"publicationDate":"2019-12-12","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":null,"resultStr":null,"platform":"Semanticscholar","paperid":"123590273","PeriodicalName":null,"FirstCategoryId":null,"ListUrlMain":null,"RegionNum":0,"RegionCategory":"","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":"","EPubDate":null,"PubModel":null,"JCR":null,"JCRName":null,"Score":null,"Total":0}
R. CésarA.Cárdenas, V. Grisales, Carlos Andrés Collazos Morales, H. Cerón-Muñoz, Paola Ariza Colpas, Roger Caputo-Llanos
{"title":"Quadrotor Modeling and a PID Control Approach","authors":"R. CésarA.Cárdenas, V. Grisales, Carlos Andrés Collazos Morales, H. Cerón-Muñoz, Paola Ariza Colpas, Roger Caputo-Llanos","doi":"10.1007/978-3-030-44689-5_25","DOIUrl":"https://doi.org/10.1007/978-3-030-44689-5_25","url":null,"abstract":"","PeriodicalId":224881,"journal":{"name":"International Conference on Intelligent Human Computer Interaction","volume":"2 1","pages":"0"},"PeriodicalIF":0.0,"publicationDate":"2019-12-12","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":null,"resultStr":null,"platform":"Semanticscholar","paperid":"129680160","PeriodicalName":null,"FirstCategoryId":null,"ListUrlMain":null,"RegionNum":0,"RegionCategory":"","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":"","EPubDate":null,"PubModel":null,"JCR":null,"JCRName":null,"Score":null,"Total":0}
Gonzalo Jiménez, Carlos Andrés Collazos Morales, Emiro de la Hoz Franco, Paola Ariza Colpas, Ramón Enrique Ramayo González, Adriana Maldonado-Franco
{"title":"Wavelet Transform Selection Method for Biological Signal Treatment","authors":"Gonzalo Jiménez, Carlos Andrés Collazos Morales, Emiro de la Hoz Franco, Paola Ariza Colpas, Ramón Enrique Ramayo González, Adriana Maldonado-Franco","doi":"10.1007/978-3-030-44689-5_3","DOIUrl":"https://doi.org/10.1007/978-3-030-44689-5_3","url":null,"abstract":"","PeriodicalId":224881,"journal":{"name":"International Conference on Intelligent Human Computer Interaction","volume":"1 1","pages":"0"},"PeriodicalIF":0.0,"publicationDate":"2019-12-12","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":null,"resultStr":null,"platform":"Semanticscholar","paperid":"130526559","PeriodicalName":null,"FirstCategoryId":null,"ListUrlMain":null,"RegionNum":0,"RegionCategory":"","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":"","EPubDate":null,"PubModel":null,"JCR":null,"JCRName":null,"Score":null,"Total":0}
Arup Kumar Pal, Dipanjan Roy, G. V. Kumar, B. Chatterjee, L. N. Sharma, A. Banerjee, C. N. Gupta
{"title":"Empirical Mode Decomposition Algorithms for Classification of Single-Channel EEG Manifesting McGurk Effect","authors":"Arup Kumar Pal, Dipanjan Roy, G. V. Kumar, B. Chatterjee, L. N. Sharma, A. Banerjee, C. N. Gupta","doi":"10.1007/978-3-030-44689-5_5","DOIUrl":"https://doi.org/10.1007/978-3-030-44689-5_5","url":null,"abstract":"","PeriodicalId":224881,"journal":{"name":"International Conference on Intelligent Human Computer Interaction","volume":"59 1","pages":"0"},"PeriodicalIF":0.0,"publicationDate":"2019-12-12","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":null,"resultStr":null,"platform":"Semanticscholar","paperid":"128851456","PeriodicalName":null,"FirstCategoryId":null,"ListUrlMain":null,"RegionNum":0,"RegionCategory":"","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":"","EPubDate":null,"PubModel":null,"JCR":null,"JCRName":null,"Score":null,"Total":0}