Assimilating human feedback from autonomous vehicle interaction in reinforcement learning models

Richard Fox,Elliot A Ludvig

doi:10.1007/s10458-024-09659-4

Abstract

A significant challenge for real-world automated vehicles (AVs) is their interaction with human pedestrians. This paper develops a methodology to directly elicit the AV behaviour pedestrians find suitable by collecting quantitative data that can be used to measure and improve an algorithm's performance. Starting with a Deep Q Network (DQN) trained on a simple Pygame/Python-based pedestrian crossing environment, the reward structure was adapted to allow adjustment by human feedback. Feedback was collected by eliciting behavioural judgements collected from people in a controlled environment. The reward was shaped by the inter-action vector, decomposed into feature aspects for relevant behaviours, thereby facilitating both implicit preference selection and explicit task discovery in tandem. Using computational RL and behavioural-science techniques, we harness a formal iterative feedback loop where the rewards were repeatedly adapted based on human behavioural judgments. Experiments were conducted with 124 participants that showed strong initial improvement in the judgement of AV behaviours with the adaptive reward structure. The results indicate that the primary avenue for enhancing vehicle behaviour lies in the predictability of its movements when introduced. More broadly, recognising AV behaviours that receive favourable human judgments can pave the way for enhanced performance.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Assimilating human feedback from autonomous vehicle interaction in reinforcement learning models

Abstract

Talk to us

Similar Papers

More From: Autonomous Agents and Multi-Agent Systems

Lead the way for us

Journal: Autonomous Agents and Multi-Agent Systems	Publication Date: Jun 26, 2024
License type: CC BY 4.0

Similar Papers

Guidance law based on deep Q network algorithm
Xianjun He ... Mingyu Wu
Journal of Physics: Conference Series | VOL. 2235
Xianjun He, et. al.Xianjun He ... Mingyu Wu
01 May 2022
Journal of Physics: Conference Series | VOL. 2235

Deep Recurrent Q-Network with Truncated History
Hyunwoo Oh ... Tomoyuki Kaneko
-
Hyunwoo Oh, et. al.Hyunwoo Oh ... Tomoyuki Kaneko
01 Nov 2018
01 Nov 2018

A Novel Approach to the Job Shop Scheduling Problem Based on the Deep Q-Network in a Cooperative Multi-Access Edge Computing Ecosystem.
Junhyung Moon ... Jongpil Jeong
Sensors (Basel, Switzerland) | VOL. 21
Junhyung Moon, et. al.Junhyung Moon ... Jongpil Jeong
02 Jul 2021
Sensors (Basel, Switzerland) | VOL. 21

Generation of Diverse Stages in Turn-Based Role-Playing Game using Reinforcement Learning
Sanggyu Nam ... Kokolo Ikeda
-
Sanggyu Nam, et. al.Sanggyu Nam ... Kokolo Ikeda
01 Aug 2019
01 Aug 2019

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Assimilating human feedback from autonomous vehicle interaction in reinforcement learning models

Abstract

Talk to us

Similar Papers

More From: Autonomous Agents and Multi-Agent Systems