Faculty Publications

Human-Agent Transfer From Observations

Bikramjit Banerjee, University of Southern MississippiFollow
Sneha Racharla, University of Southern MississippiFollow

Document Type

Article

Publication Date

1-1-2020

School

Computing Sciences and Computer Engineering

Abstract

Learning from human demonstration (LfD), among many speedup techniques for reinforcement learning (RL), has seen many successful applications. We consider one LfD technique called human–agent transfer (HAT), where a model of the human demonstrator’s decision function is induced via supervised learning and used as an initial bias for RL. Some recent work in LfD has investigated learning from observations only, that is, when only the demonstrator’s states (and not its actions) are available to the learner. Since the demonstrator’s actions are treated as labels for HAT, supervised learning becomes untenable in their absence. We adapt the idea of learning an inverse dynamics model from the data acquired by the learner’s interactions with the environment and deploy it to fill in the missing actions of the demonstrator. The resulting version of HAT—called state-only HAT (SoHAT)—is experimentally shown to preserve some advantages of HAT in benchmark domains with both discrete and continuous actions. This paper also establishes principled modifications of an existing baseline algorithm—called A3C—to create its HAT and SoHAT variants that are used in our experiments.

Publication Title

Knowledge Engineering Review

Volume

Recommended Citation

Banerjee, B., Racharla, S. (2020). Human-Agent Transfer From Observations. Knowledge Engineering Review, 36.
Available at: https://aquila.usm.edu/fac_pubs/19121

Link to Full Text

Find in your library

COinS

Faculty Publications

Human-Agent Transfer From Observations

Document Type

Publication Date

School

Abstract

Publication Title

Volume

Recommended Citation

Search

Browse

Author Corner

Faculty Publications

Human-Agent Transfer From Observations

Authors

Document Type

Publication Date

School

Abstract

Publication Title

Volume

Recommended Citation

Share

Search

Browse

Author Corner