Repository logo
Communities & Collections
All of DSpace
  • English
  • العربية
  • বাংলা
  • Català
  • Čeština
  • Deutsch
  • Ελληνικά
  • Español
  • Suomi
  • Français
  • Gàidhlig
  • हिंदी
  • Magyar
  • Italiano
  • Қазақ
  • Latviešu
  • Nederlands
  • Polski
  • Português
  • Português do Brasil
  • Srpski (lat)
  • Српски
  • Svenska
  • Türkçe
  • Yкраї́нська
  • Tiếng Việt
Log In
New user? Click here to register.Have you forgotten your password?
  1. Home
  2. Browse by Author

Browsing by Author "Hasan Rafi, Mohammad Mehdi"

Filter results by typing the first few letters
Now showing 1 - 1 of 1
  • Results Per Page
  • Sort Options
  • Thumbnail Image
    Item
    Implementation of reinforcement learning architecture to augment an AI that can self-learn to play video games
    (BRAC University, 2023-01) Mahmud, Aqil; Khan, Aswat Karim; Hasan Rafi, Mohammad Mehdi; Fahim, Kazi Rayhan; Rasel, Annajiat Alim; Khan, Rubayat Ahmed
    This paper is intended to be a practical guide in terms of getting up and running with reinforcement learning. Ideally, it aims to bridge the gap between practi cal implementation and the theories available for RL. The theory of reinforcement learning involves two main components: an environment, which is the game itself and an agent, which performs an action based on its observation from the environ ment. Initially, no in-game rules will be given to the agent and it will be rewarded or punished based on the action that it will take. The goal is to increase Proximal Policy Optimization (PPO) to maximize the reward that our agent will get, so over time it will learn what action to take in order to do so. Therefore, we will develop an AI agent that will be able to learn how to play one of the most popular arcade games of all time, Street Fighter. We preprocess our game environment and apply hyperparameter tuning using PyTorch, Stable Baselines, and Optuna to do it. This approach will basically train different types of RL architecture and find a model with the most weighted parameters. Moreover, we are going to Fine Tune that model and run our test cases on it. We are going to see how a reinforcement learning algorithm learns to play.

© Open Research Bangladesh

  • Privacy policy
  • End User Agreement
  • Send Feedback