A-TD3: An Adaptive Asynchronous Twin Delayed Deep Deterministic for Continuous Action Spaces.
Jiaolv WuQ. M. Jonathan WuShuyue ChenFarhad PourpanahDetian HuangPublished in: IEEE Access (2022)
Keyphrases
- action space
- reinforcement learning
- state space
- markov decision processes
- continuous action
- real valued
- reinforcement learning algorithms
- continuous state
- state and action spaces
- temporal difference
- stochastic processes
- continuous state spaces
- function approximators
- action selection
- control policies
- markov decision problems
- single agent
- reinforcement learning problems
- skill learning
- heuristic search
- policy iteration
- learning algorithm
- markov decision process
- reinforcement learning methods
- function approximation
- dynamic programming
- non stationary
- partially observable
- state variables
- markov chain
- optimal policy