Analyzing optimization landscape of recent policy optimization methods in deep RL

dc.contributor.advisorRashid, Warida
dc.contributor.advisorIslam, Riashat
dc.contributor.authorKhan, Mahir Asaf
dc.contributor.authorAshraf, Adib
dc.contributor.authorAmin, Tahmid Adib
dc.date.accessioned2023-05-23T04:43:23Z
dc.date.available2023-05-23T04:43:23Z
dc.date.issued2022-05
dc.descriptionCataloged from PDF version of thesis.
dc.descriptionIncludes bibliographical references (pages 42-43).
dc.descriptionThis thesis is submitted in partial fulfillment of the requirements for the degree of Bachelor of Science in Computer Science, 2022.
dc.description.abstractIn this work we will analyze control variates and baselines in policy optimization methods in deep reinforcement learning (RL). Recently there has been a lot of progress in policy gradient methods in deep RL, where baselines are typically used for variance reduction. However, there has been recent progress on the mirage of state and state-action dependent baselines in policy gradients. To this end, it is not clear how control variates play a role in the optimization landscape of policy gradients. This work will dive into understanding the landscape issues of policy optimization, to see whether control variates are only for variance reduction or whether they play a role in smoothing out the optimization landscape. Our work will further investigate the issues of different optimizers used in deep RL experiments, and ablation studies of the interplay of control variates and optimizers in policy gradients from an optimization perspective.
dc.identifier.otherID 22141075
dc.identifier.otherID 20241063
dc.identifier.otherID 22141076
dc.identifier.otherhttps://dspace.bracu.ac.bd/server/api/core/items/f01f1682-5806-48d4-a2d3-b2aeb5d549a7
dc.identifier.urihttp://hdl.handle.net/10361/18306
dc.language.isoen
dc.publisherBRAC University
dc.sourceBRAC University Institutional Repository
dc.subjectOptimization landscape
dc.subjectPolicy optimization
dc.subjectDeep reinforcement learning
dc.subjectVariance reduction
dc.subjectControl variates
dc.titleAnalyzing optimization landscape of recent policy optimization methods in deep RL
dc.typeThesis

Files

Original bundle

Now showing 1 - 1 of 1
Thumbnail Image
Name:
22141075, 20241063, 22141076_CSE.pdf
Size:
1.51 MB
Format:
Adobe Portable Document Format