Zurück zu den Ergebnissen
Bibliografischer Datensatz · Ansicht und Zugriff
Artículo de revista

Multi-objective inventory optimization using reinforcement learning: a comparative study on profitability and carbon emissions

Abdulrahman Sorour et al · Nature Portfolio · 2026

Ergänzendes Material verfügbar
Schnellübersicht. Prüfen Sie die grundlegenden Angaben und öffnen Sie den Inhalt über die Hauptschaltfläche. Die Seite zeigt nur die Informationen, die zum Identifizieren, Zitieren und Öffnen des Werks nötig sind.
Fortlaufende Publikation

3D scan-based classification of Chinese young female hand morphology

Diese fortlaufende Publikation enthält 688 zugehörige Inhalte.

Zugriff auf die Ressource

Öffnen Sie den Inhalt über die Hauptoption oder wählen Sie eine andere verfügbare Quelle.

DOAJ DOAJ Articles
Entrar por DOAJ
Hauptzugriff

Ergänzendes Material verfügbar

El enlace apunta a material asociado, anexos, tablas, datos o página complementaria. No se marca como libro/texto completo.
Material öffnen

Übersicht

Descripción general del contenido del recurso.

Abstract Inventory management is a core part of supply chains, and over the years it has been increasingly challenged by the need to balance economic performance with environmental considerations. While prior reinforcement learning (RL) studies have incorporated carbon emissions indirectly through cost penalties or regulatory constraints, this work addresses an existing gap by treating emissions as an independent optimization objective. This study examines RL as an adaptive decision‑making approach for inventory optimization with two objectives: maximizing profit and minimizing carbon emissions. The problem is formulated as a Markov Decision Process, and four RL algorithms Proximal Policy Optimization (PPO), Phasic Policy Gradient (PPG), Advantage Actor‑Critic (A2C), and Double Deep Q‑Network (DDQN) are evaluated under identical experimental conditions. Carbon emissions are explicitly modeled in the reward function rather than embedded within operating costs. The results show that PPG achieves the highest profitability with only a modest increase in emissions, while DDQN converges faster but yields lower profit overall. Sensitivity analysis indicates that reward weighting strongly influences policy behavior, with PPO providing the most stable trade‑off between profitability and emissions.

Zitieren

Elegí el formato que necesitás y copiá la referencia al portapapeles.

APA 7

al, A. S. E. (2026). Multi-objective inventory optimization using reinforcement learning: a comparative study on profitability and carbon emissions. https://doi.org/10.1038/s41598-026-44293-y

MLA

al, Abdulrahman Sorour et. "Multi-objective inventory optimization using reinforcement learning: a comparative study on profitability and carbon emissions." 2026. https://doi.org/10.1038/s41598-026-44293-y.

Chicago

al, Abdulrahman Sorour et. 2026. "Multi-objective inventory optimization using reinforcement learning: a comparative study on profitability and carbon emissions.". https://doi.org/10.1038/s41598-026-44293-y.

Harvard

al, A. S. E. 2026, Multi-objective inventory optimization using reinforcement learning: a comparative study on profitability and carbon emissions, Nature Portfolio, available at: https://doi.org/10.1038/s41598-026-44293-y [Accessed 6 Aug. 2026].

Teilen und drucken

Speichern Sie den Datensatz, kopieren Sie den Permalink oder drucken Sie ihn als PDF.

Referenz exportieren

Exportieren Sie den Datensatz in gängigen Formaten für Literaturverwaltungsprogramme.

Ressourcendetails

Bibliografische Angaben zur Prüfung, ob es sich um das richtige Material handelt.

Titel
Multi-objective inventory optimization using reinforcement learning: a comparative study on profitability and carbon emissions
Autor / Mitwirkende
Abdulrahman Sorour et al
Verlag
Nature Portfolio
Erscheinungsjahr
2026
ISSN
2045-2322
ISSN
2045-2322
Sprache
Inglés

Schlagwörter

Entdecken Sie über diese Schlagwörter weitere verwandte Ressourcen.

Kopiert