Ada-NAV: Adaptive Trajectory Length-Based Sample Efficient Policy Learning for Robotic Navigation

Patel, Bhrij; Weerakoon, Kasun; Suttle, Wesley A.; Koppel, Alec; Sadler, Brian M.; Zhou, Tianyi; Bedi, Amrit Singh; Manocha, Dinesh

Computer Science > Robotics

arXiv:2306.06192 (cs)

[Submitted on 9 Jun 2023 (v1), last revised 20 Mar 2024 (this version, v5)]

Title:Ada-NAV: Adaptive Trajectory Length-Based Sample Efficient Policy Learning for Robotic Navigation

Authors:Bhrij Patel, Kasun Weerakoon, Wesley A. Suttle, Alec Koppel, Brian M. Sadler, Tianyi Zhou, Amrit Singh Bedi, Dinesh Manocha

View PDF HTML (experimental)

Abstract:Trajectory length stands as a crucial hyperparameter within reinforcement learning (RL) algorithms, significantly contributing to the sample inefficiency in robotics applications. Motivated by the pivotal role trajectory length plays in the training process, we introduce Ada-NAV, a novel adaptive trajectory length scheme designed to enhance the training sample efficiency of RL algorithms in robotic navigation tasks. Unlike traditional approaches that treat trajectory length as a fixed hyperparameter, we propose to dynamically adjust it based on the entropy of the underlying navigation policy. Interestingly, Ada-NAV can be applied to both existing on-policy and off-policy RL methods, which we demonstrate by empirically validating its efficacy on three popular RL methods: REINFORCE, Proximal Policy Optimization (PPO), and Soft Actor-Critic (SAC). We demonstrate through simulated and real-world robotic experiments that Ada-NAV outperforms conventional methods that employ constant or randomly sampled trajectory lengths. Specifically, for a fixed sample budget, Ada-NAV achieves an 18\% increase in navigation success rate, a 20-38\% reduction in navigation path length, and a 9.32\% decrease in elevation costs. Furthermore, we showcase the versatility of Ada-NAV by integrating it with the Clearpath Husky robot, illustrating its applicability in complex outdoor environments.

Comments:	11 pages, 9 figures, 2 tables
Subjects:	Robotics (cs.RO); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
Cite as:	arXiv:2306.06192 [cs.RO]
	(or arXiv:2306.06192v5 [cs.RO] for this version)
	https://doi.org/10.48550/arXiv.2306.06192

Submission history

From: Bhrij Patel [view email]
[v1] Fri, 9 Jun 2023 18:45:15 UTC (40,320 KB)
[v2] Fri, 29 Sep 2023 13:04:00 UTC (49,332 KB)
[v3] Mon, 2 Oct 2023 21:40:07 UTC (49,332 KB)
[v4] Tue, 19 Mar 2024 16:16:16 UTC (8,078 KB)
[v5] Wed, 20 Mar 2024 17:36:07 UTC (8,078 KB)

Computer Science > Robotics

Title:Ada-NAV: Adaptive Trajectory Length-Based Sample Efficient Policy Learning for Robotic Navigation

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Robotics

Title:Ada-NAV: Adaptive Trajectory Length-Based Sample Efficient Policy Learning for Robotic Navigation

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators