Double Q($σ$) and Q($σ, λ$): Unifying Reinforcement Learning Control Algorithms | Read Paper on Bytez