Repository logo
  • English
  • Català
  • Čeština
  • Deutsch
  • Español
  • Français
  • Gàidhlig
  • Italiano
  • Latviešu
  • Magyar
  • Nederlands
  • Polski
  • Português
  • Português do Brasil
  • Srpski (lat)
  • Suomi
  • Svenska
  • Türkçe
  • Tiếng Việt
  • Қазақ
  • বাংলা
  • हिंदी
  • Ελληνικά
  • Српски
  • Yкраї́нська
  • Log In
    New user? Click here to register. Have you forgotten your password?
Repository logo
  • Colleges, Institutes & Collections
  • Browse AAU-ETD
  • English
  • Català
  • Čeština
  • Deutsch
  • Español
  • Français
  • Gàidhlig
  • Italiano
  • Latviešu
  • Magyar
  • Nederlands
  • Polski
  • Português
  • Português do Brasil
  • Srpski (lat)
  • Suomi
  • Svenska
  • Türkçe
  • Tiếng Việt
  • Қазақ
  • বাংলা
  • हिंदी
  • Ελληνικά
  • Српски
  • Yкраї́нська
  • Log In
    New user? Click here to register. Have you forgotten your password?
  1. Home
  2. Browse by Author

Browsing by Author "Migbar Abera Shibru"

Now showing 1 - 1 of 1
Results Per Page
Sort Options
  • No Thumbnail Available
    Item
    CGAWM - Curriculum-Guided Adversarial World Model for Data-Efficient Adversarial Reinforcement Learning
    (Addis Ababa University, 2026-02) Migbar Abera Shibru; Beakal Gizachew; Natnael Argaw (Co-Advisor)
    Deep Reinforcement Learning (DRL) has achieved remarkable success in single-agent domains; however, Multi-Agent Reinforcement Learning (MARL) remains plagued by severe sample inefficiency and training instability, particularly in competitive zero-sum games. The non-stationarity inherent in adversarial environments where the opponent’s strategy evolves continuously often prevents agents from converging to optimal policies, leading to cycling behaviors or catastrophic forgetting. To address these challenges, this thesis proposes the Curriculum-Guided Adversarial World Model (CGAWM), a framework that integrates three methodological pillars: an Adversarial Dynamics Model (ADeM) that amplifies data efficiency via synthetic rollouts, an Adversarial Curriculum that stabilizes training by progressively increasing opponent competence, and an Uncertainty-Guided Exploration mechanism that incentivizes the agent to investigate high-entropy states. Evaluated on the PettingZoo simple adversary benchmark, experimental results demonstrate that CGAWM achieves a 136% improvement in asymptotic reward compared to state-of-the-art Model-Free PPO baselines. Furthermore, the proposed method exhibited superior reliability, achieving a consistentconvergence rate across random seeds where baseline methods failed to solve the task. Qualitative analysis confirmed the emergence of sophisticated game-theoretic behaviors, including Split-Coverage and Defensive Hovering, indicating that the agent successfully developed a primitive Theory of Mind. These findings suggest that integrating world models with structured curricula provides a scalable path toward robust autonomous systems in adversarial domains.

Home |Privacy policy |End User Agreement |Send Feedback |Library Website

Addis Ababa University © 2023