Autonomous Assessment of Demonstration Sufficiency via Bayesian Inverse Reinforcement Learning

Trinh, Tu; Chen, Haoyu; Brown, Daniel S.

doi:10.1145/3610977.3634984

Computer Science > Machine Learning

arXiv:2211.15542 (cs)

[Submitted on 28 Nov 2022 (v1), last revised 2 Jan 2024 (this version, v3)]

Title:Autonomous Assessment of Demonstration Sufficiency via Bayesian Inverse Reinforcement Learning

Authors:Tu Trinh, Haoyu Chen, Daniel S. Brown

View PDF HTML (experimental)

Abstract:We examine the problem of determining demonstration sufficiency: how can a robot self-assess whether it has received enough demonstrations from an expert to ensure a desired level of performance? To address this problem, we propose a novel self-assessment approach based on Bayesian inverse reinforcement learning and value-at-risk, enabling learning-from-demonstration ("LfD") robots to compute high-confidence bounds on their performance and use these bounds to determine when they have a sufficient number of demonstrations. We propose and evaluate two definitions of sufficiency: (1) normalized expected value difference, which measures regret with respect to the human's unobserved reward function, and (2) percent improvement over a baseline policy. We demonstrate how to formulate high-confidence bounds on both of these metrics. We evaluate our approach in simulation for both discrete and continuous state-space domains and illustrate the feasibility of developing a robotic system that can accurately evaluate demonstration sufficiency. We also show that the robot can utilize active learning in asking for demonstrations from specific states which results in fewer demos needed for the robot to still maintain high confidence in its policy. Finally, via a user study, we show that our approach successfully enables robots to perform at users' desired performance levels, without needing too many or perfectly optimal demonstrations.

Comments:	Prior version appears in proceedings of AAAI FSS-22 Symposium "Lessons Learned for Autonomous Assessment of Machine Abilities (LLAAMA)". Current version appears in proceedings of HRI '24, March 11-14, 2024, Boulder, CO, USA
Subjects:	Machine Learning (cs.LG)
Cite as:	arXiv:2211.15542 [cs.LG]
	(or arXiv:2211.15542v3 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2211.15542
Related DOI:	https://doi.org/10.1145/3610977.3634984

Submission history

From: Tu Trinh [view email]
[v1] Mon, 28 Nov 2022 16:48:24 UTC (5,416 KB)
[v2] Tue, 29 Nov 2022 21:53:24 UTC (5,416 KB)
[v3] Tue, 2 Jan 2024 06:36:38 UTC (4,232 KB)

Computer Science > Machine Learning

Title:Autonomous Assessment of Demonstration Sufficiency via Bayesian Inverse Reinforcement Learning

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Autonomous Assessment of Demonstration Sufficiency via Bayesian Inverse Reinforcement Learning

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators