Using reinforcement learning techniques to solve continuous‐time non‐linear optimal tracking problem without system dynamics. Issue 12 (1st August 2016)