Relative importance sampling for off-policy actor-critic in deep reinforcement learning
Mahammad Humayoo(Anshan Normal University), Xueqi Cheng(Institute of Computing Technology), Naveed Ur Rehman Junejo(Anshan Normal University), Xiaoqing Dong(Anshan Normal University), Liming Miao(Shanghai Academy of Agricultural Sciences), Shuwei Qiu(East Tennessee State University), Zakir Ullah(University of Chinese Academy of Sciences), Peitao Wang(Anshan Normal University), Gengzhong Zheng(Hanshan Normal University), Zexun Zhou(Anshan Normal University)
Cited by 3
Related Papers
Highly Compressible Integrated Supercapacitor–Piezoresistance‐Sensor System with CNT–PDMS Sponge for Health Monitoring
|Small|2017|311
High efficiency power management and charge boosting strategy for a triboelectric nanogenerator
|Nano Energy|2017|222