Dynamically constructed (PO)MDPs for adaptive robot planning

19Citations
Citations of this article
39Readers
Mendeley users who have this article in their library.
Get full text

Abstract

To operate in human-robot coexisting environments, intelligent robots need to simultaneously reason with commonsense knowledge and plan under uncertainty. Markov decision processes (MDPs) and partially observable MDPs (POMDPs), are good at planning under uncertainty toward maximizing long-term rewards; P-LOG, a declarative programming language under Answer Set semantics, is strong in commonsense reasoning. In this paper, we present a novel algorithm called iCORPP to dynamically reason about, and construct (PO)MDPs using P-LOG. iCORPP successfully shields exogenous domain attributes from (PO)MDPs, which limits computational complexity and enables (PO)MDPs to adapt to the value changes these attributes produce. We conduct a number of experimental trials using two example problems in simulation and demonstrate iCORPP on a real robot. Results show significant improvements compared to competitive baselines.

Cite

CITATION STYLE

APA

Zhang, S., Khandelwal, P., & Stone, P. (2017). Dynamically constructed (PO)MDPs for adaptive robot planning. In 31st AAAI Conference on Artificial Intelligence, AAAI 2017 (pp. 3855–3862). AAAI press. https://doi.org/10.1609/aaai.v31i1.11042

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free