Diagnosing Non-Intermittent Anomalies in Reinforcement Learning Policy Executions (Short Paper)
Testo / risorsa
Testo / risorsa
Due to the safety risks and training sample inefficiency, it is often preferred to develop controllers in simulation. However, minor differences between the simulation and the real world can cause a significant sim-to-re...
Idioma Inglés
Página del recurso disponiblePágina de referencia del recurso. El texto completo no está confirmado automáticamente.
Página del recurso