期刊论文

期刊论文详细信息

Journal of the Brazilian Computer Society
Compulsory Flow Q-Learning: an RL algorithm for robot navigation based on partial-policy and macro-states

Costa, Anna Helena Reali¹ Universidade de São Paulo, São Paulo, Brasil¹ Silva, Valdinei Freire da¹
关键词: machine learning; reinforcement learning; abstraction; partial-policy; macro-states.;
DOI : 10.1007/BF03194507
学科分类：农业科学（综合）
来源: Springer U K
PDF

【摘要】

Reinforcement Learning is carried out on-line, through trial-and-error interactions of the agent with the environment, which can be very time consuming when considering robots. In this paper we contribute a new learning algorithm, CFQ-Learning, which uses macro-states, a low-resolution discretisation of the state space, and a partial-policy to get around obstacles, both of them based on the complexity of the environment structure. The use of macro-states avoids convergence of algorithms, but can accelerate the learning process. In the other hand, partial-policies can guarantee that an agent fulfils its task, even through macro-state. Experiments show that the CFQ-Learning performs a good balance between policy quality and learning rate.

【授权许可】

Unknown

【预览】

附件列表
Files	Size	Format	View
RO201912010163974ZK.pdf	1115KB	PDF	download


	文献评价指标
	下载次数：9次	浏览次数：16次

京公网安备340104078870146号 878987797 028-85220240

OAinOne平台基于对开放资源的发现、遴选和评价方式，发现、获取、集成9类优质的开放科技资源，包括开放期刊、开放会议论文、开放课件、科技政策、开放学位论文、开放图书、开放科技报告、科研项目、开放科学数据。同时，为实现开放知识资源普遍服务、个性化服务、精准服务，基于OAinONE集成的丰富开放资源，开发建设领域开放知识资源服务定制工具(OAtoYOU)、开放资源评价评估体系（OAEvaluation），建设集成OAinONE资源及其他第三方资源的OA Hub，及其面向我院分布式大数据知识资源系统及其他第三方的开放接口服务，并打造特色专题数据库产品建设，包括科技政策集成及趋势平台、开放课程大讲堂等。此外，OAinOne构建开放知识资源建设的可持续发展机制，支持我院研究所特色馆藏资源、自建资源、古籍资源等在OAinONE平台上的集成、开放、共享。

【 摘 要 】

【 授权许可】

【 预 览 】

【摘要】

【授权许可】

【预览】