Torna ai risultati
Scheda bibliografica · Consultazione e accesso
Artículo

Traffic expertise meets residual RL: Knowledge-informed model-based residual reinforcement learning for CAV trajectory control

Zihao Sheng et al · Tsinghua University Press · 2024

Accesso aperto disponibile
Lettura rapida. Controlla i dati essenziali della risorsa e accedi al contenuto con il pulsante principale. La scheda mostra solo le informazioni necessarie per identificare, citare e aprire l’opera.

Accesso alla risorsa

Apri il contenuto dall’opzione principale o scegli un’altra fonte disponibile.

DOAJ DOAJ Articles
Entrar por DOAJ
Accesso principale

Accesso aperto disponibile

Recurso identificado como acceso abierto, sin confirmar automáticamente si es texto completo directo.
Apri risorsa

Riepilogo

Descripción general del contenido del recurso.

Model-based reinforcement learning (RL) is anticipated to exhibit higher sample efficiency than model-free RL by utilizing a virtual environment model. However, obtaining sufficiently accurate representations of environmental dynamics is challenging because of uncertainties in complex systems and environments. An inaccurate environment model may degrade the sample efficiency and performance of model-based RL. Furthermore, while model-based RL can improve sample efficiency, it often still requires substantial training time to learn from scratch, potentially limiting its advantages over model-free approaches. To address these challenges, this paper introduces a knowledge-informed model-based residual reinforcement learning framework aimed at enhancing learning efficiency by infusing established expert knowledge into the learning process and avoiding the issue of beginning from zero. Our approach integrates traffic expert knowledge into a virtual environment model, employing the intelligent driver model (IDM) for basic dynamics and neural networks for residual dynamics, thus ensuring adaptability to complex scenarios. We propose a novel strategy that combines traditional control methods with residual RL, facilitating efficient learning and policy optimization without the need to learn from scratch. The proposed approach is applied to connected automated vehicle (CAV) trajectory control tasks for the dissipation of stop-and-go waves in mixed traffic flows. The experimental results demonstrate that our proposed approach enables the CAV agent to achieve superior performance in trajectory control compared with the baseline agents in terms of sample efficiency, traffic flow smoothness and traffic mobility.

Come citare

Elegí el formato que necesitás y copiá la referencia al portapapeles.

APA 7

al, Z. S. E. (2024). Traffic expertise meets residual RL: Knowledge-informed model-based residual reinforcement learning for CAV trajectory control. https://doi.org/10.1016/j.commtr.2024.100142

MLA

al, Zihao Sheng et. "Traffic expertise meets residual RL: Knowledge-informed model-based residual reinforcement learning for CAV trajectory control." 2024. https://doi.org/10.1016/j.commtr.2024.100142.

Chicago

al, Zihao Sheng et. 2024. "Traffic expertise meets residual RL: Knowledge-informed model-based residual reinforcement learning for CAV trajectory control.". https://doi.org/10.1016/j.commtr.2024.100142.

Harvard

al, Z. S. E. 2024, Traffic expertise meets residual RL: Knowledge-informed model-based residual reinforcement learning for CAV trajectory control, Tsinghua University Press, available at: https://doi.org/10.1016/j.commtr.2024.100142 [Accessed 6 Aug. 2026].

Condividi e stampa

Salva la scheda, copia il link permanente o stampala in PDF.

Esporta riferimento

Esporta il record nei formati più comuni per usarlo con un gestore bibliografico.

Dettagli della risorsa

Informazioni bibliografiche utili per verificare che sia il materiale corretto.

Titolo
Traffic expertise meets residual RL: Knowledge-informed model-based residual reinforcement learning for CAV trajectory control
Autore / collaboratori
Zihao Sheng et al
Editore
Tsinghua University Press
Anno di pubblicazione
2024
ISSN
2772-4247
ISSN
2772-4247
Lingua
Inglés

Soggetti

Esplora risorse correlate a partire da questi soggetti.

Copiato