Anda belum login :: 27 Nov 2024 17:17 WIB
Detail
ArtikelSequence-Learning Algorithm Based on Backward Chaining  
Oleh: Joshi, Sanjay S. ; Guilhabert, Benoit
Jenis: Article from Journal - e-Journal
Dalam koleksi: Adaptive Behavior vol. 14 no. 1 (Mar. 2006), page 53–71.
Topik: sequence-learning; chaining; trial-and-error; stochastic; animal-training; learning automaton
Fulltext: 53.pdf (481.07KB)
Isi artikelThis article considers the problem of learning the correct temporal sequence of discrete behaviors from a finite behavior set that will lead to completion of a complex task, using only stochastic reinforcement from the environment. A trial-and-error learning algorithm is proposed that is inspired by backward chaining from the animal training discipline. The procedure is analytically formulated using a serial composition of finite action-set learning automata with delay. Simulation of the proposed algorithm shows that the algorithm does indeed lead to sequence learning. The effect of parametric variation in the magnitude and quality of reinforcement is investigated in both theory and simulation. It is shown that a fundamental trade off exists between quality and speed of learning. It is also shown that the algorithm has the ability to learn desirable action sequences among several feasible action sequences through the use of relative rewards, which may be interpreted using the Bellman principle of optimality.
Opini AndaKlik untuk menuliskan opini Anda tentang koleksi ini!

Kembali
design
 
Process time: 0.015625 second(s)