Back to results
Bibliographic record · Consultation and access
Artículo de revista

Multimodal sentiment analysis: hybrid classification model with image and text feature descriptors

P. Vasanthi et al · Nature Portfolio · 2026

Open access available
Quick overview. Review the resource’s basic details, then access the content using the main button. This page shows only the information needed to identify, cite, and open the work.
Serial publication

3D scan-based classification of Chinese young female hand morphology

This serial publication contains 688 related contents.

Resource access

Open the content from the main option or choose another available source.

DOAJ DOAJ Articles
Entrar por DOAJ
Main access

Open access available

Recurso identificado como acceso abierto, sin confirmar automáticamente si es texto completo directo.
Open resource

Summary

Descripción general del contenido del recurso.

Abstract Understanding human emotions across multiple modalities such as text and images, is increasingly important for applications including content personalization, social media analysis, and Human–Computer Interaction (HCI). Conventional sentiment analysis methods often rely on a single modality, overlooking complementary information from other sources. This paper proposes a novel multimodal sentiment analysis framework that integrates text and image data. Text preprocessing includes tokenization, stopword removal, and stemming, while image preprocessing employs object detection. From the preprocessed text, N-grams, emojis, and Normalized Dispersion Coefficient (NDC)-based Term frequency-inverse document frequency (TF-IDF) features are extracted. Then, the improved multitexon and Shape Local Binary Texture (SLBT) features are derived from the preprocessed images. A hybrid sentiment analysis model is introduced, combining an optimized Deep Maxout and a Modified Sigmoid (MS)-based Bidirectional Gated Recurrent Unit (MS-Bi-GRU) models. Here, the extracted text features are subjected to an optimized Deep Maxout model; on the other hand, the MS-based Bi-GRU model trains the image features. Specifically, the transfer learning strategy is employed in the MS-based Bi-GRU model which leveraging the knowledge of a pre-trained model to efficiently train the feature set to improve efficiency. The Deep Maxout weights are optimized using a novel Innovative Beluga Whale Optimization Algorithm (IBwOA). Based on the outcomes from both models, the final results will be determined. Experimental results demonstrate that the proposed approach outperforms conventional models across multiple performance metrics.

How to cite

Elegí el formato que necesitás y copiá la referencia al portapapeles.

APA 7

al, P. V. E. (2026). Multimodal sentiment analysis: hybrid classification model with image and text feature descriptors. https://doi.org/10.1038/s41598-026-42912-2

MLA

al, P. Vasanthi et. "Multimodal sentiment analysis: hybrid classification model with image and text feature descriptors." 2026. https://doi.org/10.1038/s41598-026-42912-2.

Chicago

al, P. Vasanthi et. 2026. "Multimodal sentiment analysis: hybrid classification model with image and text feature descriptors.". https://doi.org/10.1038/s41598-026-42912-2.

Harvard

al, P. V. E. 2026, Multimodal sentiment analysis: hybrid classification model with image and text feature descriptors, Nature Portfolio, available at: https://doi.org/10.1038/s41598-026-42912-2 [Accessed 7 Aug. 2026].

Share and print

Save the record, copy its permanent link, or print it as a PDF.

Export reference

You can export the record in common formats for use in a reference manager.

Resource details

Bibliographic information to help confirm that this is the correct material.

Title
Multimodal sentiment analysis: hybrid classification model with image and text feature descriptors
Author / contributors
P. Vasanthi et al
Publisher
Nature Portfolio
Publication year
2026
ISSN
2045-2322
ISSN
2045-2322
Language
English

Subjects

Explore related resources through these subjects.

Copied