Reinforcement Learning and Optimal Control

Reinforcement Learning and Optimal Control

评分 6.9 分
格式EPUB
ISBN9781886529397
出版社
语言英文

内容简介

This book considers large and challenging multistage decision problems, which can be solved in principle by dynamic programming, but their exact solution is computationally intractable. It can be used as a textbook or for self-study in conjunction with instructional videos and slides, and other supporting material, which are available from the author's website. The book discusses solution methods that rely on approximations to produce suboptimal policies with adequate performance. These methods are known by several essentially equivalent names: reinforcement learning, approximate dynamic programming, and neuro-dynamic programming. They underlie, among others, the recent impressive successes of self-learning in the context of games such as chess and Go. One of the aims of the book is to explore the common boundary between artificial intelligence and optimal control, and to form a bridge that is accessible by workers with background in either field. Another aim is to organize coherently the broad mosaic of methods that have proved successful in practice while having a solid theoretical and/or logical foundation. This may help researchers and practitioners to find their way through the maze of competing ideas that constitute the current state of the art. The mathematical style of this book is somewhat different than other books by the same author. While we provide a rigorous, albeit short, mathematical account of the theory of finite and infinite horizon dynamic programming, and some fundamental approximation methods, we rely more on intuitive explanations and less on proof-based insights. We also illustrate the methodology with many example algorithms and applications. Dimitri Bertsekas is McAffee Professor of Electrical Engineering and Computer Science at the Massachusetts Institute of Technology, and a member of the National Academy of Engineering. He has researched a broad variety of subjects from optimization theory, control theory, parallel and distributed comput
立即下载

点击查看全部下载链接(含网盘地址及提取码)

相关书籍

At This Defining Moment: Barack Obama's Presidential Candidacy and the New Politics of Race
At This Defining Moment: Barack Obama's Presidential Candidacy and the New Politics of Race
At This Defining Moment: Barack Obama's Presidential Candidacy and the New Politics of RaceEnid Lynette Logan
Intelligent Transport Systems and Travel Behaviour: 13th Scientific and Technical Conference "Transport Systems. Theory and Practice 2016" Selected Papers
Intelligent Transport Systems and Travel Behaviour: 13th Scientific and Technical Conference "Transport Systems. Theory and Practice 2016" Selected Papers
Intelligent Transport Systems and Travel Behaviour: 13th Scientific and Technical Conference "Transport Systems. Theory and Practice 2016" Selected PapersGrzegorz Sierpiński
Bright Unequivocal Eye: Poems, Papers, and Remembrances from the First Jane Kenyon Conference
Bright Unequivocal Eye: Poems, Papers, and Remembrances from the First Jane Kenyon Conference
Bright Unequivocal Eye: Poems, Papers, and Remembrances from the First Jane Kenyon ConferenceBert G. Hornback
"It's the Pictures That Got Small": Charles Brackett on Billy Wilder and Hollywood's Golden Age
"It's the Pictures That Got Small": Charles Brackett on Billy Wilder and Hollywood's Golden Age
"It's the Pictures That Got Small": Charles Brackett on Billy Wilder and Hollywood's Golden AgeCharles Brackett, Anthony Slide
AI销冠:一个人顶一个团队的销售术
AI销冠:一个人顶一个团队的销售术
AI销冠:一个人顶一个团队的销售术唐兴通
人性博弈
人性博弈
人性博弈李尚龙
正念指导
正念指导
正念指导莉兹·霍尔
有效地招聘
有效地招聘
有效地招聘保罗·法尔科内
Isis and Sarapis in the Roman World
Isis and Sarapis in the Roman World
Isis and Sarapis in the Roman WorldSaroltaA.Takacs
ICD-10
ICD-10
ICD-10
LA CEREMONIA DEL ADIOS
LA CEREMONIA DEL ADIOS
LA CEREMONIA DEL ADIOSSimonedeBeauvoir
Love, Money and Obligation
Love, Money and Obligation
Love, Money and ObligationPatcharinLapanun