Improving News Recommendation Accuracy Through Multimodal Variational Autoencoder and Adversarial Training
| dc.contributor.author | Tang, Pei | |
| dc.contributor.author | Zhu, Shuo | |
| dc.contributor.author | Alatas, Bilal | |
| dc.date.accessioned | 2026-08-12T17:26:43Z | |
| dc.date.issued | 2025 | |
| dc.department | Fırat Üniversitesi | |
| dc.description.abstract | Traditional recommender systems, such as collaborative filtering and content filtering, have inherent limitations, including cold start issues and challenges in filtering information. Additionally, the sparsity and noise in news data significantly affect recommendation accuracy. To address these issues, this paper introduces a multi-head self-attention mechanism to capture both long-term and local dependencies within user sequences. A Multimodal Variational Autoencoder (MVAVE) is constructed based on the multi-head self-attention mechanism, incorporating noise injection for denoising to mitigate the impact of data noise on recommendation performance. Building upon MVAVE, an adversarial multimodal data sequence recommender system is further developed by integrating MVAVE with a Generative Adversarial Network (GAN). Through adversarial training, the system reduces the reconstruction loss of user-item interactions and enhances recommendation accuracy. Moreover, multimodal data fusion is utilized to provide richer and more comprehensive information. Experimental results demonstrate superior recommendation performance on the NewsStories and Amazon datasets, achieving optimal results when the number of multi-head attention heads (HEAD) is set to 6 and the hyperparameter is 0.2. Specifically, on the NewsStories dataset, the system achieves a Recall of 21.74% and a Precision of 22.06%; on the Amazon dataset, it achieves a Recall of 18.67% and a Precision of 18.36%. This system effectively captures user preferences, avoids recommending irrelevant content, resolves the limitations of traditional recommender systems, and significantly improves recommendation accuracy. | |
| dc.description.sponsorship | Startup Fund for Advanced Talents of Putian University [2023054.00]; Fujian Provincial Department of Ethnic and Religious Affairs 2024 General Project; Chinese national community [FJMZ202408]; Fujian Provincial Social Science Fund Special Commissioned Youth Project [FJ2024TWQX002] | |
| dc.description.sponsorship | This work received the support of ''Startup Fund for Advanced Talents of Putian University: Research on the meaning production ofmultimodal graphic notes of Meizhou Island tourism in the age of social media'', the project number is 2023054.00. The support of''Fujian Provincial Department of Ethnic and Religious Affairs 2024 General Project: Research on the development of the discoursesystem of Mazu short video stories from the perspective of the Chinese national community'', the project number is FJMZ202408. And''Fujian Provincial Social Science Fund Special Commissioned Youth Project: Research on short video production and optimization pathof Fujian's regional image from the perspective of international communication'', the project number is FJ2024TWQX002. | |
| dc.identifier.doi | 10.1109/ACCESS.2025.3568514 | |
| dc.identifier.endpage | 85278 | |
| dc.identifier.issn | 2169-3536 | |
| dc.identifier.orcid | 0000-0002-3513-0329 | |
| dc.identifier.scopus | 2-s2.0-105004814054 | |
| dc.identifier.scopusquality | Q1 | |
| dc.identifier.startpage | 85269 | |
| dc.identifier.uri | https://doi.org/10.1109/ACCESS.2025.3568514 | |
| dc.identifier.uri | https://hdl.handle.net/11508/54938 | |
| dc.identifier.volume | 13 | |
| dc.identifier.wos | WOS:001492129400048 | |
| dc.identifier.wosquality | Q2 | |
| dc.indekslendigikaynak | Web of Science | |
| dc.indekslendigikaynak | Scopus | |
| dc.language.iso | en | |
| dc.publisher | Ieee-Inst Electrical Electronics Engineers Inc | |
| dc.relation.ispartof | Ieee Access | |
| dc.relation.publicationcategory | Makale - Uluslararası Hakemli Dergi - Kurum Öğretim Elemanı | |
| dc.rights | info:eu-repo/semantics/openAccess | |
| dc.snmz | KA_WoS_20260511 | |
| dc.subject | Recommender systems | |
| dc.subject | Autoencoders | |
| dc.subject | Accuracy | |
| dc.subject | Generative adversarial networks | |
| dc.subject | Noise | |
| dc.subject | Training | |
| dc.subject | Feature extraction | |
| dc.subject | Data integration | |
| dc.subject | Visualization | |
| dc.subject | Knowledge graphs | |
| dc.subject | Variational autoencoder | |
| dc.subject | generative adversarial networks | |
| dc.subject | multimodal recommender systems | |
| dc.subject | multi-head attention mechanismmeta-learning | |
| dc.subject | deep learning | |
| dc.subject | convolutional neural networks | |
| dc.subject | neural networks | |
| dc.subject | computer vision | |
| dc.title | Improving News Recommendation Accuracy Through Multimodal Variational Autoencoder and Adversarial Training | |
| dc.type | Article |







