Unveiling the Significance of Toddler-Inspired Reward Transition in Goal-Oriented Reinforcement Learning.
Junseok ParkYoonsung KimHee bin YooMin Whoo LeeKibeom KimWon-Seok ChoiMinsu LeeByoung-Tak ZhangPublished in: AAAI (2024)
Keyphrases
- goal oriented
- reinforcement learning
- transition model
- function approximation
- state space
- reward function
- eligibility traces
- requirements analysis
- reinforcement learning algorithms
- requirements engineering
- optimal policy
- temporal difference
- model free
- learning algorithm
- optimal control
- markov decision processes
- partially observable environments
- process oriented
- data marts
- total reward
- dynamic programming
- partially observable
- markov decision process
- reinforcement learning methods
- multi agent
- case study
- learning process
- learning agent
- markov decision problems
- decentralized control
- inverse reinforcement learning
- data mining
- action selection