Reinforcement Learning for Sequential Decision and Optimal Control

ISBN: 9789811977831
Код товара 153635

Shengbo Eben Li / Ли

Reinforcement Learning for Sequential Decision and Optimal Control

Reinforcement Learning for Sequential Decision and Optimal Control

ISBN: 9789811977831
Код товара 153635

XML_ID: 15811281

Нет в наличии

Когда книга появится на складе, напишем вам на почту
  • Автор

    Shengbo Eben Li / Ли

  • Издатель

    Springer

  • Тип обложки

    Hardback

  • Размеры

    174 x 247 x 34

  • Год издания

    2022

  • Вес (г)

    328

  • ISBN

    9789811977831

  • Язык

    ENG

  • Кол-во страниц

    472

О чём книга?

Have you ever wondered how AlphaZero learns to defeat the top human Go players' Do you have any clues about how an autonomous driving system can gradually develop self-driving skills beyond normal drivers' What is the key that enables AlphaStar to make decisions in Starcraft, a notoriously difficult strategy game that has partial information and complex rules' The core mechanism underlying those recent technical breakthroughs is reinforcement learning (RL), a theory that can help an agent to develop the self-evolution ability through continuing environment interactions. In the past few years, the AI community has witnessed phenomenal success of reinforcement learning in various fields, including chess games, computer games and robotic control. RL is also considered to be a promising and powerful tool to create general artificial intelligence in the future. As an interdisciplinary field of trial-and-error learning and optimal control, RL resembles how humans reinforce their intelligence by interacting with the environment and provides a principled solution for sequential decision making and optimal control in large-scale and complex problems. Since RL contains a wide range of new concepts and theories, scholars may be plagued by a number of questions: What is the inherent mechanism of reinforcement learning' What is the internal connection between RL and optimal control' How has RL evolved in the past few decades, and what are the milestones' How do we choose and implement practical and effective RL algorithms for real-world scenarios' What are the key challenges that RL faces today, and how can we solve them' What is the current trend of RL research' You can find answers to all those questions in this book. The purpose of the book is to help researchers and practitioners take a comprehensive view of RL and understand the in-depth connection between RL and optimal control. The book includes not only systematic and thorough explanations of theoretical basics but also methodical guidance of practical algorithm implementations. The book intends to provide a comprehensive coverage of both classic theories and recent achievements, and the content is carefully and logically organized, including basic topics such as the main concepts and terminologies of RL, Markov decision process (MDP), Bellman’s optimality condition, Monte Carlo learning, temporal difference learning, stochastic dynamic programming, function approximation, policy gradient methods, approximate dynamic programming, and deep RL, as well as the latest advances in action and state constraints, safety guarantee, reference harmonization, robust RL, partially observable MDP, multiagent RL, inverse RL, offline RL, and so on.

Chapter 1 Introduction of Reinforcement Learning.- Chapter 2 Principles of RL Problems.- Chapter 3 Model-free Indirect RL: Monte Carlo.- Chapter 4 Model-Free Indirect RL: Temporal-Difference.- Chapter 5 Model-based Indirect RL: Dynamic Programming.- Chapt

Отзывы

К сожалению, отзывов пока нет.
Похожие книги
Around the table
изд. 2026

Mcgee, Shea Around the table

  • изд. 2026
5 491 ₽
Функциональная магнитно-резонансная томография головного мозга
изд. 2026

Паникратова Я.Р., Лебедева И.С., Горбунов А.В., Горев В.В., Чайка Ю.А. Функциональная магнитно-резонансная томография головного мозга

  • RUS
  • изд. 2026
Функциональная магнитно-резонансная томография головного мозга / Я.Р. Паникратова, И.С. Лебедева, А.В. Горбунов, В.В. Горев, Ю.А. Чайка. - М.: Логосфера, 2026. - 224 с.; 15,6 см. ISBN 978-5-98657-125-6 Книга посвящена одному из ключевых методов нейровизуализации - функциональной магнитно-резонансно...
1 990 ₽
Nursing education in the ai era
изд. 2026

Nursing education in the ai era

  • изд. 2026
This book provides a groundbreaking approach to integrating artificial intelligence (AI) in ways that enhance learning, bridge the gap between theory and clinical practice, and ultimately prepare learners to deliver safe, equitable, and high-quality patient care. With structured, practical guidance ...
7 927 ₽
TNM: классификация злокачественных опухолей
изд. 2026

Дж.Д. Брайерли и др. (ред.) TNM: классификация злокачественных опухолей

  • RUS
  • изд. 2026
TNM: Классификация злокачественных опухолей / Под ред. Дж.Д. Брайерли и др.; пер. с англ. и научн. ред. Е.А. Дубовой, О.Р. Катуниной, К.А. Павлова. 3-е изд. на русском языке. - М.: Логосфера, 2026. - 424 с. : 14,0 см. ISBN 978-598657-124-9 Девятое издание TNM: Классификация злокачественных опухолей...
1 990 ₽
Circuits as Systems
изд. 2026

Robert W. Erickson / Роберт В. Эриксон Circuits as Systems

  • ENG
  • изд. 2026
This groundbreaking textbook provides coverage for the second semester, core course in electronic circuits. Unlike most textbooks for this course, this one covers the mathematics of frequency-domain analysis, the traditional language of electrical engineering, in the context of real engineering appl...
11 948 ₽
Рекомендации Европейского общества кардиологов. Повышенное артериальное давление и гипертензия
изд. 2025

Рекомендации Европейского общества кардиологов. Повышенное артериальное давление и гипертензия

  • RUS
  • изд. 2025
В данных Рекомендациях обобщены и оценены доказательства по повышенному артериальному давлению и артериальной гипертензии. Приведены современные классификации, рекомендации по лечению на основе последних доказательных данных, уточнена роль генетики, даны рекомендации по ведению не только пациентов с...