00
Days
00
Hrs
00
Min
00
Sec
Submit Your Paper

Detecting Misinformation Using Multimodal AI Models on Social Media Platforms

Authors

Ashwini Sonawane

Department of Computer Science, Dr. D. Y. Patil Arts, Commerce and Science College, Pimpri, Pune, Maharashtra, India (IN)

Sayali Shinde

Department of Computer Science, Dr. D. Y. Patil Arts, Commerce and Science College, Pimpri, Pune, Maharashtra, India (IN)

Article Information

DOI: 10.51583/IJLTEMAS.2025.1413SP002

Subject Category: Computer Science

Volume/Issue: 14/13 | Page No: 7-10

Publication Timeline

Submitted: 2025-10-22

Published: 2025-10-22

Abstract

Abstract: Misinformation on social media has become a critical challenge, impacting public opinion, health, and democracy. Traditional text-based methods for misinformation detection often fall short because social media content is increasingly multimodal, containing images, videos, and text. This paper explores the use of multimodal AI models that integrate visual, textual, and contextual features to improve the accuracy of misinformation detection on social media platforms. We present an overview of recent advancements, propose a multimodal framework, and discuss experimental results, challenges, and future research directions.

Keywords

Multimodal Fusion, Natural Language Processing, Multimodal AI, Social Network Analysis, Deepfake Detection

Downloads

References

1. Aronoff, S. (1989). Geographic Information Systems: A Management Perspective. Ottawa: WDL Publications. [Google Scholar] [Crossref]

2. Jin, Z., Cao, J., Guo, H., Zhang, Y., & Luo, J. (2017). Multimodal fusion with recurrent neural networks for rumor detection on microblogs. ACM Multimedia. [Google Scholar] [Crossref]

3. Kiela, D., Bulat, L., & Clark, S. (2019). Learning Multimodal Representations with Sparse Attention. arXiv preprint arXiv:1902.00751. [Google Scholar] [Crossref]

4. Devlin, J., Chang, M. W., Lee, K., & Toutanova, K. (2019). BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. NAACL. [Google Scholar] [Crossref]

5. Lu, J., Batra, D., Parikh, D., & Lee, S. (2019). VilBERT: Pretraining Task-Agnostic Visiolinguistic Representations for Vision-and-Language Tasks. NeurIPS. [Google Scholar] [Crossref]

6. Shu, K., Sliva, A., Wang, S., Tang, J., & Liu, H. (2017). Fake News Detection on Social Media: A Data Mining Perspective. ACM SIGKDD Explorations. [Google Scholar] [Crossref]

7. Wang, Y., Ma, F., Jin, Z., Yuan, Y., Xun, G., Jha, K., & Gao, J. (2018). EANN: Event Adversarial Neural Networks for Multi-Modal Fake News Detection. KDD. [Google Scholar] [Crossref]

Metrics

Views & Downloads

Similar Articles

© 2026 IJLTEMAS · RSIS International. All rights reserved. ISSN 2278-2540.