مقاله یادگیری تقویتی مبتنی بر منیفولد از طریق بازسازی محلی به صورت خطی  
پروژه 24

پروژه 24

دانلود پروژه و حل تمرین و گزارش کارآموزی و تحقیق و مقاله و جزوه و کتاب

نظرسنجی سایت

رشته تحصیلی شما؟

اشتراک در خبرنامه

جهت عضویت در خبرنامه لطفا ایمیل خود را ثبت نمائید

Captcha

آمار بازدید

  • بازدید امروز : 424
  • بازدید دیروز : 1115
  • بازدید کل : 3337421

مقاله یادگیری تقویتی مبتنی بر منیفولد از طریق بازسازی محلی به صورت خطی


مقاله یادگیری تقویتی مبتنی بر منیفولد از طریق بازسازی محلی به صورت خطی

عنوان مقاله فارسی: یادگیری تقویتی مبتنی بر منیفولد از طریق بازسازی محلی به صورت خطی

عنوان مقاله لاتین: Manifold-Based Reinforcement Learning via Locally Linear Reconstruction

نویسندگان: Xin Xu; Zhenhua Huang; Lei Zuo; Haibo He

تعداد صفحات: 13

سال انتشار: 2017

زبان: لاتین


Abstract:

Feature representation is critical not only for pattern recognition tasks but also for reinforcement learning (RL) methods to solve learning control problems under uncertainties. In this paper, a manifold-based RL approach using the principle of locally linear reconstruction (LLR) is proposed for Markov decision processes with large or continuous state spaces. In the proposed approach, an LLR-based feature learning scheme is developed for value function approximation in RL, where a set of smooth feature vectors is generated by preserving the local approximation properties of neighboring points in the original state space. By using the proposed feature learning scheme, an LLR-based approximate policy iteration (API) algorithm is designed for learning control problems with large or continuous state spaces. The relationship between the value approximation error of a new data point and the estimated values of its nearest neighbors is analyzed. In order to compare different feature representation and learning approaches for RL, a comprehensive simulation and experimental study was conducted on three benchmark learning control problems. It is illustrated that under a wide range of parameter settings, the LLR-based API algorithm can obtain better learning control performance than the previous API methods with different feature representation schemes.

  انتشار : ۱۰ خرداد ۱۴۰۰               تعداد بازدید : 669

برچسب های مهم

http://kia-ir.ir

در صورت هرگونه مشکل و مغایرت در دانلود فایل ها به پشتیبانی سایت مراجعه کنید

فید خبر خوان    نقشه سایت    تماس با ما