The Role of Attention Mechanism in Generating Image Captions: An Innovative Approach with Neural Network-Based Seq2seq Model

dc.contributor.authorKaraca, Zeynep
dc.contributor.authorDaş, Bihter
dc.date.accessioned2026-08-12T16:07:49Z
dc.date.issued2024
dc.departmentFırat Üniversitesi
dc.description.abstractThis study addresses important contributions to generating text from images, aiming to create meaning in various fields such as entertainment, communication, commerce, security, and education by establishing a connection between visual and textual content. This process aims to increase the accessibility, understandability, and processability of content by converting image data into meaningful text. Therefore, advances and studies in this field are extremely important. This study focuses on the effect of the combination of deep neural network models and attention mechanisms in creating more meaningful captions from images. Experiments performed on the Flickr8k dataset highlight the abilities of Seq2seq and VGG19 models to generate titles compatible with reference sentences. By using the dynamic focusing feature of the attention mechanism, the model effectively captures detailed aspects of images. The findings of this study have the potential to push the boundaries of multimodal data processing and representation with the effective integration of visual and textual information by adding information that the attention mechanism works more effectively together with the Seq2seq model. © 2024, Sakarya University. All rights reserved.
dc.identifier.doi10.35377/saucis...1339931
dc.identifier.endpage102
dc.identifier.issn2636-8129
dc.identifier.issue1
dc.identifier.scopus2-s2.0-85214399882
dc.identifier.scopusqualityQ3
dc.identifier.startpage92
dc.identifier.trdizinid1233749
dc.identifier.urihttps://doi.org/10.35377/saucis...1339931
dc.identifier.urihttps://search.trdizin.gov.tr/tr/yayin/detay/1233749
dc.identifier.urihttps://hdl.handle.net/11508/40913
dc.identifier.volume7
dc.indekslendigikaynakScopus
dc.indekslendigikaynakTR-Dizin
dc.language.isoen
dc.publisherSakarya University
dc.relation.ispartofSakarya University Journal of Computer and Information Sciences
dc.relation.publicationcategoryMakale - Uluslararası Hakemli Dergi - Kurum Öğretim Elemanı
dc.rightsinfo:eu-repo/semantics/openAccess
dc.snmzKA_Scopus_20260511
dc.subjectAttention mechanism; Deep learning; Image Capturing; Image-to-text; Seq2seq model; VGG19
dc.titleThe Role of Attention Mechanism in Generating Image Captions: An Innovative Approach with Neural Network-Based Seq2seq Model
dc.typeArticle

Dosyalar