Back to results
Bibliographic record · Consultation and access
Artículo de revista

Hierarchical multi-agent reinforcement learning for retrieval-augmented industrial document question answering

Yihong Qian et al · Nature Portfolio · 2026

Supplementary material available
Quick overview. Review the resource’s basic details, then access the content using the main button. This page shows only the information needed to identify, cite, and open the work.
Serial publication

3D scan-based classification of Chinese young female hand morphology

This serial publication contains 688 related contents.

Resource access

Open the content from the main option or choose another available source.

DOAJ DOAJ Articles
Entrar por DOAJ
Main access

Supplementary material available

El enlace apunta a material asociado, anexos, tablas, datos o página complementaria. No se marca como libro/texto completo.
Open material

Summary

Descripción general del contenido del recurso.

Abstract Multimodal industrial documents–such as operation manuals, circuit diagrams, and parameter tables–contain domain knowledge distributed across text, images, and document layout. However, most existing retrieval-augmented generation (RAG) frameworks rely on static retrieval and fusion policies with fixed modality weights and uniform retrieval depth, making them less adaptable to diverse query intents and dynamic cross-modal dependencies. As a result, they often retrieve incomplete evidence and yield suboptimal reasoning in complex long-document scenarios. To address these challenges, we propose MARL-RAGDoc, a hierarchical multi-agent reinforcement learning framework for multimodal retrieval-augmented reasoning. A high-level coordinator agent dynamically allocates modality weights and retrieval depth based on query characteristics, while specialized text, image, and table agents perform fine-grained evidence selection within their respective candidate pools. A collaborative reasoning module integrates the retrieved evidence and provides hierarchical reward signals to continuously optimize retrieval policies. Experimental results on multiple multimodal document benchmarks demonstrate that MARL-RAGDoc consistently outperforms baselines in both retrieval accuracy and reasoning performance, while remaining computationally efficient. Our code and dataset are publicly available at https://github.com/Yihong-Q/MARL-RAGDoc .

How to cite

Elegí el formato que necesitás y copiá la referencia al portapapeles.

APA 7

al, Y. Q. E. (2026). Hierarchical multi-agent reinforcement learning for retrieval-augmented industrial document question answering. https://doi.org/10.1038/s41598-026-41684-z

MLA

al, Yihong Qian et. "Hierarchical multi-agent reinforcement learning for retrieval-augmented industrial document question answering." 2026. https://doi.org/10.1038/s41598-026-41684-z.

Chicago

al, Yihong Qian et. 2026. "Hierarchical multi-agent reinforcement learning for retrieval-augmented industrial document question answering.". https://doi.org/10.1038/s41598-026-41684-z.

Harvard

al, Y. Q. E. 2026, Hierarchical multi-agent reinforcement learning for retrieval-augmented industrial document question answering, Nature Portfolio, available at: https://doi.org/10.1038/s41598-026-41684-z [Accessed 8 Aug. 2026].

Share and print

Save the record, copy its permanent link, or print it as a PDF.

Export reference

You can export the record in common formats for use in a reference manager.

Resource details

Bibliographic information to help confirm that this is the correct material.

Title
Hierarchical multi-agent reinforcement learning for retrieval-augmented industrial document question answering
Author / contributors
Yihong Qian et al
Publisher
Nature Portfolio
Publication year
2026
ISSN
2045-2322
ISSN
2045-2322
Language
English

Subjects

Explore related resources through these subjects.

Copied