Self-corrective Apprenticeship Learning for Quadrotor Control

Master Thesis (2019)
Author(s)

Haoran Yuan (TU Delft - Aerospace Engineering)

Contributor(s)

Erik-jan van Kampen – Mentor (Control & Simulation)

Q. P. Chu – Graduation committee member (Control & Simulation)

Dominic Dirkx – Graduation committee member (TU Delft - Aerospace Engineering)

Bo Sun – Graduation committee member (Control & Simulation)

Faculty
Aerospace Engineering
More Info
expand_more
Publication Year
2019
Language
English
Graduation Date
27-11-2019
Awarding Institution
Delft University of Technology
Programme
Aerospace Engineering
Faculty
Aerospace Engineering
Downloads counter
248
Collections
thesis
Reuse Rights

Other than for strictly personal use, it is not permitted to download, forward or distribute the text or part of it, without the consent of the author(s) and/or copyright holder(s), unless the work is under an open content license such as Creative Commons.

Abstract

The control of aircraft can be carried out by Reinforcement Learning agents; however, the difficulty of obtaining sufficient training samples often makes this approach infeasible. Demonstrations can be used to facilitate the learning process, yet algorithms such as Apprenticeship Learning generally fail to produce a policy that outperforms the demonstrator, and thus cannot efficiently generate policies. In this paper, a model-free learning algorithm with Reinforcement Learning in the loop, based on Apprenticeship Learning, is therefore proposed. This algorithm uses external measurement to improve on the initial demonstration, finally producing a policy that surpasses the demonstration. Efficiency is further improved by utilising the policies produced during the learning process. The empirical results for simulated quadrotor control show that the proposed algorithm is effective and can even learn good policies from a bad demonstration.

Files

MScThesis_HaoranYuan.pdf
(pdf | 3.21 Mb)
License info not available