株式会社極東書店トップ商品一覧Transfer Learning for Multiagent Reinforcement Learning Systems.

商品詳細

Transfer Learning for Multiagent Reinforcement Learning Systems.

Transfer Learning for Multiagent Reinforcement Learning Systems.

・ISBN 978-3-031-00463-6 paper EUR 59.99

¥16,034.- (税込) (※)価格はご注文時の参考価格となります。
納品価格につきましては書籍の入荷時点で確定となります。
版元の原価改定、外国為替の変動等により異なる場合がございますので、予めご了承下さい。

お気に入り
著者・編者Leno da Silva, Felipe / Costa, Anna Helena Reali,
シリーズ (Synthesis Lectures on Artificial Intelligence and Machine Learning)
出版社 (Springer International Publishing AG, SZ)
出版年月2021
ページ数111 pp.
言語ENG
ニュース番号<A02-98341>

解説

Learning to solve sequential decision-making tasks is difficult. Humans take years exploring the environment essentially in a random way until they are able to reason, solve difficult tasks, and collaborate with other humans towards a common goal. Artificial Intelligent agents are like humans in this aspect. Reinforcement Learning (RL) is a well-known technique to train autonomous agents through interactions with the environment. Unfortunately, the learning process has a high sample complexity to infer an effective actuation policy, especially when multiple agents are simultaneously actuating in the environment.

However, previous knowledge can be leveraged to accelerate learning and enable solving harder tasks. In the same way humans build skills and reuse them by relating different tasks, RL agents might reuse knowledge from previously solved tasks and from the exchange of knowledge with other agents in the environment. In fact, virtually all of the most challenging tasks currently solved by RL rely on embedded knowledge reuse techniques, such as Imitation Learning, Learning from Demonstration, and Curriculum Learning.

This book surveys the literature on knowledge reuse in multiagent RL. The authors define a unifying taxonomy of state-of-the-art solutions for reusing knowledge, providing a comprehensive discussion of recent progress in the area. In this book, readers will find a comprehensive discussion of the many ways in which knowledge can be reused in multiagent sequential decision-making tasks, as well as in which scenarios each of the approaches is more efficient. The authors also provide their view of the current low-hanging fruit developments of the area, as well as the still-open big questions that could result in breakthrough developments. Finally, the book provides resources to researchers who intend to join this area or leverage those techniques, including a list of conferences, journals, and implementation tools.

This book will be useful for a wide audience; and will hopefully promote new dialogues across communities and novel developments in the area.