Facebook pixel tracking

The Science and Information (SAI) Organization publishes open-access peer-reviewed journals in computer science and artificial intelligence.

Contact Info
Website thesai.org
Follow Us
Contact Info
Follow Us
Research Article | Open Access |

Multimodal Cognitive Mapping Framework for Context-Aware Figurative Language Understanding

Author 1: R. Swathi Gudipati Author 2: Neena PC Author 3: K. Ezhilmathi Author 4: M. Durairaj Author 5: S. Farhad Author 6: Elangovan Muniyandy Author 7: Padmashree V
International Journal of Advanced Computer Science and Applications (IJACSA) · Vol. 16, No. 11 · Published 2025

DOI: https://doi.org/10.14569/IJACSA.2025.0161164

Abstract

Learning figurative language, including idioms, metaphors, and similes, remains challenging due to subtle cultural, contextual, and multimodal cues that cannot be inferred from literal meanings alone. Traditional unimodal and text-only approaches, such as CLS-BERT, LaBSE, and mUSE, often fail to capture these deeper semantic patterns, resulting in reduced accuracy and limited cultural generalization. This study introduces a context-aware multimodal learning framework that integrates textual embeddings from a Graph-Enhanced Transformer (HCGT) with visual embeddings from CLIP, fused through a graph-based cross-modal attention mechanism, and refined using a cognitive mapping layer. This architecture models human-like semantic reasoning by aligning literal and figurative senses across modalities while maintaining conceptual structure through graph-driven representation learning. Experiments conducted on idiom, metaphor, simile, and multimodal meme datasets include preprocessing steps such as text cleaning, tokenization, image normalization, and label standardization. The framework achieves an accuracy of 90%, surpassing state-of-the-art text-only transformer baselines by 3–4%. Explainable AI tools, including attention heatmaps and SHAP values, validate the interpretability of the model by highlighting influential textual tokens and visual regions. The results confirm that integrating multimodal embeddings with cognitive mapping substantially enhances performance, interpretability, and cultural sensitivity in figurative language understanding.

Keywords

How to Cite this Article

Gudipati, R. S., PC, N., Ezhilmathi, K., Durairaj, M., Farhad, S., Muniyandy, E., & V, P. (2025). Multimodal Cognitive Mapping Framework for Context-Aware Figurative Language Understanding. International Journal of Advanced Computer Science and Applications, 16(11). https://doi.org/10.14569/IJACSA.2025.0161164

Gudipati, R. Swathi, et al.. "Multimodal Cognitive Mapping Framework for Context-Aware Figurative Language Understanding." International Journal of Advanced Computer Science and Applications, vol. 16, no. 11, 2025, https://doi.org/10.14569/IJACSA.2025.0161164.

@article{Gudipati2025,
  title     = {Multimodal Cognitive Mapping Framework for Context-Aware Figurative Language Understanding},
  journal   = {International Journal of Advanced Computer Science and Applications},
  volume    = {16},
  number    = {11},
  year      = {2025},
  publisher = {The Science and Information Organization},
  author    = {R. Swathi Gudipati and Neena PC and K. Ezhilmathi and M. Durairaj and S. Farhad and Elangovan Muniyandy and Padmashree V},
  doi       = {10.14569/IJACSA.2025.0161164},
  url       = {https://doi.org/10.14569/IJACSA.2025.0161164}
}

Open Access — licensed under a Creative Commons Attribution 4.0 International License. Unrestricted use, distribution, and reproduction in any medium, even commercially, as long as the original work is properly cited.