Deep Reinforcement Learning for List-wise Recommendations.
Xiangyu ZhaoLiang ZhangZhuoye DingDawei YinYihong ZhaoJiliang TangPublished in: CoRR (2018)
Keyphrases
- reinforcement learning
- recommender systems
- function approximation
- pairwise
- state space
- robotic control
- reinforcement learning algorithms
- model free
- deep learning
- dynamic programming
- markov decision processes
- reinforcement learning methods
- learning algorithm
- learning process
- web search
- optimal policy
- dynamical systems
- ranked list
- multi agent
- user feedback
- recommendation systems
- robot control