Experimental analysis of eligibility traces strategies in temporal difference learning. (16th December 2008)
- Record Type:
- Journal Article
- Title:
- Experimental analysis of eligibility traces strategies in temporal difference learning. (16th December 2008)
- Main Title:
- Experimental analysis of eligibility traces strategies in temporal difference learning
- Authors:
- Leng, Jinsong
Jain, Lakhmi
Fyfe, Colin - Abstract:
- Temporal difference (TD) learning is a model-free reinforcement learning technique, which adopts an infinite horizon discount model and uses an incremental learning technique for dynamic programming. The state value function is updated in terms of sample episodes. Utilising eligibility traces is a key mechanism in enhancing the rate of convergence. TD(λ) represents the use of eligibility traces by introducing the parameter λ. However, the underlying mechanism of eligibility traces with an approximation function has not been well understood, either from theoretical point of view or from practical point of view. The TD(λ) method has been proved to be convergent with local tabular state representation. Unfortunately, proving convergence of TD(λ) with function approximation is still an important open theoretical question. This paper aims to investigate the convergence and the effects of different eligibility traces. In this paper, we adopt Sarsa(λ) learning control algorithm with a large, stochastic and dynamic simulation environment called SoccerBots. The state value function is represented by a linear approximation function known as tile coding. The performance metrics generated from the simulation system can be used to analyse the mechanism of eligibility traces.
- Is Part Of:
- International journal of knowledge engineering and soft data paradigms. Volume 1:Number 1(2009)
- Journal:
- International journal of knowledge engineering and soft data paradigms
- Issue:
- Volume 1:Number 1(2009)
- Issue Display:
- Volume 1, Issue 1 (2009)
- Year:
- 2009
- Volume:
- 1
- Issue:
- 1
- Issue Sort Value:
- 2009-0001-0001-0000
- Page Start:
- 26
- Page End:
- 39
- Publication Date:
- 2008-12-16
- Subjects:
- agents -- temporal difference learning -- eligibility traces -- decision making -- reinforcement learning -- infinite horizon discount -- incremental learning -- dynamic simulation -- soccer games -- SoccerBots
Soft computing -- Periodicals
Statistics -- Periodicals
Information science -- Periodicals
003.05 - Journal URLs:
- http://www.inderscience.com/jhome.php?jcode=ijkesdp ↗
http://www.inderscience.com/ ↗ - Languages:
- English
- ISSNs:
- 1755-3210
- Deposit Type:
- Legaldeposit
- View Content:
- Available online (eLD content is only available in our Reading Rooms) ↗
- Physical Locations:
- British Library DSC - BLDSS-3PM
British Library STI - ELD Digital store - Ingest File:
- 8736.xml