Higher Level Application of ADP: A Next Phase for the Control Field?

Two distinguishing features of humanlike control vis-a-vis current technological control are the ability to make use of experience while selecting a control policy for distinct situations and the ability to do so faster and faster as more experience is gained (in contrast to current technological im...

Ausführliche Beschreibung

Gespeichert in:

Bibliographische Detailangaben
Veröffentlicht in:	IEEE transactions on cybernetics 2008-08, Vol.38 (4), p.901-912
1. Verfasser:	Lendaris, G.G.
Format:	Artikel
Sprache:	eng
Schlagworte:	Adaptive control Algorithm design and analysis Algorithms Approximate dynamic programming (ADP) Approximation Artificial Intelligence artificial intelligence (AI) Artificial neural networks Biomimetics - methods context context discernment Control systems Cybernetics Design Design engineering Dynamic programming experience-based identification and control (EBIC) Feedback Humans Learning neural networks (NNs) On-line systems Optimal control Optimization Programming, Linear Reinforcement reinforcement learning (RL) Studies System identification system identification (SID) Systems Theory
Online-Zugang:	Volltext bestellen
Tags:	Tag hinzufügen Keine Tags, Fügen Sie den ersten Tag hinzu!

Beschreibung
Zusammenfassung:	Two distinguishing features of humanlike control vis-a-vis current technological control are the ability to make use of experience while selecting a control policy for distinct situations and the ability to do so faster and faster as more experience is gained (in contrast to current technological implementations that slow down as more knowledge is stored). The notions of context and context discernment are important to understanding this human ability. Whereas methods known as adaptive control and learning control focus on modifying the design of a controller as changes in context occur, experience-based (EB) control entails selecting a previously designed controller that is appropriate to the current situation. Developing the EB approach entails a shift of the technologist's focus ldquoup a levelrdquo away from designing individual (optimal) controllers to that of developing online algorithms that efficiently and effectively select designs from a repository of existing controller solutions. A key component of the notions presented here is that of higher level learning algorithm. This is a new application of reinforcement learning and, in particular, approximate dynamic programming, with its focus shifted to the posited higher level, and is employed, with very promising results. The author's hope for this paper is to inspire and guide future work in this promising area.
ISSN:	1083-4419 2168-2267 1941-0492 2168-2275
DOI:	10.1109/TSMCB.2008.918073