Óbudai Egyetem Digitális Archívum
    • magyar
    • English
  • English 
    • magyar
    • English
  • Login
View Item 
  •   DSpace Home
  • 5. Folyóiratcikkek
  • Acta Polytechnica Hungarica
  • 2. 2024
  • 2.08. 2024 Volume 21, Issue No. 4.
  • View Item
  •   DSpace Home
  • 5. Folyóiratcikkek
  • Acta Polytechnica Hungarica
  • 2. 2024
  • 2.08. 2024 Volume 21, Issue No. 4.
  • View Item
JavaScript is disabled for your browser. Some features of this site may not work without it.

Improving Multiagent Actor-Critic Architectures, with Opponent Approximation and Dropout for Control

Thumbnail
View/Open
Paczolay_Harmati_144.pdf (438.3Kb)
Metadata
Show full item record
URI
http://hdl.handle.net/20.500.14044/33467
Collections
  • 2.08. 2024 Volume 21, Issue No. 4. [16]
Abstract
In the domain of reinforcement learning, solution proposals to multiagent problems are evolving. We propose a new algorithm, MADDPGX, to handle the problem of higher uncertainty created by other agents’ actions by an enemy actor approximator, and we investigate the most efficient techniques of estimations. This approximation works using a neural network, which has the input of the state and the output as the action (probably preferred by the enemy agent). We also experimented with dropout, a tool commonly used for neural networks, but has not been used efficiently for reinforcement learning until now. We have also found that in multiagent actor-critic scenarios, it can improve overall performance. Generally, our contribution is the use of action approximation of adversaries and the dropout usage in actor-critic systems, with a conclusion that the newly proposed methods will perform better in zero-sum multi-agent robot system scenarios. The experiments were conducted in a multiagent predator-prey environment.
Title
Improving Multiagent Actor-Critic Architectures, with Opponent Approximation and Dropout for Control
Author
Paczolay, Gabor
Harmati, Istvan
xmlui.dri2xhtml.METS-1.0.item-date-issued
2024
xmlui.dri2xhtml.METS-1.0.item-rights-access
Open access
xmlui.dri2xhtml.METS-1.0.item-identifier-issn
1785-8860
xmlui.dri2xhtml.METS-1.0.item-language
en
xmlui.dri2xhtml.METS-1.0.item-format-page
20 p.
xmlui.dri2xhtml.METS-1.0.item-subject-oszkar
reinforcement learning, multiagent learning, dropout, MADDPG, MADDPGX
xmlui.dri2xhtml.METS-1.0.item-description-version
Kiadói változat
xmlui.dri2xhtml.METS-1.0.item-identifiers
DOI: 10.12700/APH.21.4.2024.4.13
xmlui.dri2xhtml.METS-1.0.item-other-containerTitle
Acta Polytechnica Hungarica
xmlui.dri2xhtml.METS-1.0.item-other-containerPeriodicalYear
2024
xmlui.dri2xhtml.METS-1.0.item-other-containerPeriodicalVolume
21. évf.
xmlui.dri2xhtml.METS-1.0.item-other-containerPeriodicalNumber
4. sz.
xmlui.dri2xhtml.METS-1.0.item-type-type
Tudományos cikk
xmlui.dri2xhtml.METS-1.0.item-subject-area
Műszaki tudományok - multidiszciplináris műszaki tudományok
xmlui.dri2xhtml.METS-1.0.item-publisher-university
Óbudai Egyetem

DSpace software copyright © 2002-2016  DuraSpace
Contact Us | Send Feedback
Theme by 
Atmire NV
 

 

Browse

All of DSpaceCommunities & CollectionsBy Issue DateAuthorsTitlesSubjectsThis CollectionBy Issue DateAuthorsTitlesSubjects

My Account

LoginRegister

DSpace software copyright © 2002-2016  DuraSpace
Contact Us | Send Feedback
Theme by 
Atmire NV