Full text

Turn on search term navigation

Copyright © 2022 Bratislav Predić et al. This is an open access article distributed under the Creative Commons Attribution License (the “License”), which permits unrestricted use, distribution, and reproduction in any medium, provided the original work is properly cited. Notwithstanding the ProQuest Terms and Conditions, you may use this content in accordance with the terms of the License. https://creativecommons.org/licenses/by/4.0/

Abstract

This paper is dedicated to machine learning, the branches of machine learning, which include the methods for solving this issue, and the practical implementation of the solution to the automatic image description generation. Automatic image caption generation is one of the frequent goals of computer vision. Image description generation models must solve a larger number of complex problems to have this task successfully solved. The objects in the image must be detected and recognized, after which a logical and syntactically correct textual description is generated. For that reason, description generation is a complex problem. It is an extremely important challenge for machine learning algorithms because it represents an impersonation of a complicated human ability to encapsulate huge amounts of highlighted visual pieces of information in descriptive language. The results of the generated descriptions are compared depending on the used pretrained convolutional networks. The BLEU metrics are used to calculate the quality of the image description. Although the solution to the problem of image description automatic generation does provide us with good results, there is yet room for improvement since there are images that are not adequately described.

Details

Title
Automatic Image Caption Generation Based on Some Machine Learning Algorithms
Author
Predić, Bratislav 1 ; Manić, Daša 1 ; Saračević, Muzafer 2   VIAFID ORCID Logo  ; Karabašević, Darjan 3 ; Stanujkić, Dragiša 4 

 Faculty of Electronic Engineering, University of Niš, Aleksandra Medvedeva 14, Niš 18000, Serbia 
 Department of Computer Sciences, University of Novi Pazar, Dimitrija Tucovića bb, 36300, Novi Pazar, Serbia 
 Faculty of Applied Management, Economics and Finance, University Business Academy in Novi Sad, Belgrade, Serbia, Jevrejska 24, Belgrade 11000, Serbia 
 Technical Faculty in Bor, University of Belgrade, Vojske Jugoslavije 12, Bor 19210, Serbia 
Editor
Bogdan Smolka
Publication year
2022
Publication date
2022
Publisher
John Wiley & Sons, Inc.
ISSN
1024123X
e-ISSN
15635147
Source type
Scholarly Journal
Language of publication
English
ProQuest document ID
2653906783
Copyright
Copyright © 2022 Bratislav Predić et al. This is an open access article distributed under the Creative Commons Attribution License (the “License”), which permits unrestricted use, distribution, and reproduction in any medium, provided the original work is properly cited. Notwithstanding the ProQuest Terms and Conditions, you may use this content in accordance with the terms of the License. https://creativecommons.org/licenses/by/4.0/