Zurück zu den Ergebnissen
Bibliografischer Datensatz · Ansicht und Zugriff
Artículo

Quantifying Modality Contributions via Disentangling Multimodal Representations

Padegal Amit et al · LibraryPress@UF · 2026

Open-Access-Volltext
Schnellübersicht. Prüfen Sie die grundlegenden Angaben und öffnen Sie den Inhalt über die Hauptschaltfläche. Die Seite zeigt nur die Informationen, die zum Identifizieren, Zitieren und Öffnen des Werks nötig sind.

Zugriff auf die Ressource

Öffnen Sie den Inhalt über die Hauptoption oder wählen Sie eine andere verfügbare Quelle.

DOAJ DOAJ Articles
Entrar por DOAJ
Hauptzugriff

Open-Access-Volltext

Texto completo identificado como acceso abierto.
Text öffnen

Übersicht

Descripción general del contenido del recurso.

Quantifying modality contributions in Vision-Language Models (VLMs) remains challenging. Existing approaches rely on perturbation or gradient-based methods, which conflate inherent modality informativeness with model-specific biases and fail to capture complex cross-modal interactions. We address this gap by introducing an information-theoretic framework based on Partial Information Decomposition (PID) that decomposes internal representations into unique, redundant, and synergistic components. Our method operates directly on internal embeddings and derives an inference-only modality contribution metric from unique information scores. Applying our framework to six modern VLMs across six benchmarks, we uncover a persistent imbalance in modality contributions driven by low cross-modal synergy. Analysis reveals that fusion architecture significantly impacts the distribution of unique, redundant, and synergistic information. Our framework provides a scalable diagnostic tool for understanding and improving multimodal integration in vision-language systems.

Zitieren

Elegí el formato que necesitás y copiá la referencia al portapapeles.

APA 7

al, P. A. E. (2026). Quantifying Modality Contributions via Disentangling Multimodal Representations. https://journals.flvc.org/FLAIRS/article/view/141869

MLA

al, Padegal Amit et. "Quantifying Modality Contributions via Disentangling Multimodal Representations." 2026. https://journals.flvc.org/FLAIRS/article/view/141869.

Chicago

al, Padegal Amit et. 2026. "Quantifying Modality Contributions via Disentangling Multimodal Representations.". https://journals.flvc.org/FLAIRS/article/view/141869.

Harvard

al, P. A. E. 2026, Quantifying Modality Contributions via Disentangling Multimodal Representations, LibraryPress@UF, available at: https://journals.flvc.org/FLAIRS/article/view/141869 [Accessed 8 Aug. 2026].

Teilen und drucken

Speichern Sie den Datensatz, kopieren Sie den Permalink oder drucken Sie ihn als PDF.

Referenz exportieren

Exportieren Sie den Datensatz in gängigen Formaten für Literaturverwaltungsprogramme.

Ressourcendetails

Bibliografische Angaben zur Prüfung, ob es sich um das richtige Material handelt.

Titel
Quantifying Modality Contributions via Disentangling Multimodal Representations
Autor / Mitwirkende
Padegal Amit et al
Verlag
LibraryPress@UF
Erscheinungsjahr
2026
ISSN
2334-0754
ISSN
2334-0754
Sprache
Inglés
Kopiert