Actively Learning Gaussian Process Dynamical Systems Through Global and Local Explorations

Tang, Shengbing; Fujimoto, Kenji; Maruta, Ichiro

このアイテムのアクセス数: 108

http://hdl.handle.net/2433/278980

このアイテムのファイル:

ファイル	記述	サイズ	フォーマット
ACCESS.2022.3154095.pdf		1.73 MB	Adobe PDF	見る/開く

タイトル:	Actively Learning Gaussian Process Dynamical Systems Through Global and Local Explorations
著者:	Tang, Shengbing Fujimoto, Kenji https://orcid.org/0000-0001-6345-4884 (unconfirmed) Maruta, Ichiro https://orcid.org/0000-0002-2246-3570 (unconfirmed)
著者名の別形:	藤本, 健治丸田, 一郎
キーワード:	Active learning dynamical system Gaussian process global and local explorations
発行日:	2022
出版者:	Institute of Electrical and Electronics Engineers (IEEE)
誌名:	IEEE Access
巻:	10
開始ページ:	24215
終了ページ:	24231
抄録:	Usually learning dynamical systems by data-driven methods requires large amount of training data, which may be time consuming and expensive. Active learning, which aims at choosing the most informative samples to make learning more efficient is a promising way to solve this issue. However, actively learning dynamical systems is difficult since it is not possible to arbitrarily sample the state-action space under the constraint of system dynamics. The state-of-the-art methods for actively learning dynamical systems iteratively search for an informative state-action pair by maximizing the differential entropy of the predictive distribution, or iteratively search for a long informative trajectory by maximizing the sum of predictive variances along the trajectory. These methods suffer from low efficiency or high computational complexity and memory demand. To solve these problems, this paper proposes novel and more sample-efficient methods which combine global and local explorations. As the global exploration, the agent searches for a relatively short informative trajectory in the whole state-action space of the dynamical system. Then, as the local exploration, an action sequence is optimized to drive the system’s state towards the initial state of the local informative trajectory found by the global exploration and the agent explores this local informative trajectory. Compared to the state-of-the-art methods, the proposed methods are capable of exploring the state-action space more efficiently, and have much lower computational complexity and memory demand. With the state-of-the-art methods as baselines, the advantages of the proposed methods are verified via various numerical examples.
著作権等:	This work is licensed under a Creative Commons Attribution-NonCommercial-NoDerivatives 4.0 License.
URI:	http://hdl.handle.net/2433/278980
DOI(出版社版):	10.1109/ACCESS.2022.3154095
出現コレクション:	学術雑誌掲載論文等

アイテムの詳細レコードを表示する

Export to RefWorks

このアイテムは次のライセンスが設定されています: クリエイティブ・コモンズ・ライセンス