Asynchronous, Option-Based Multi-Agent Policy Gradient: A Conditional Reasoning Approach

March 29, 2022 · Declared Dead · 🏛 IEEE/RJS International Conference on Intelligent RObots and Systems

"No code URL or promise found in abstract"

Evidence collected by the PWNC Scanner

Authors Xubo Lyu, Amin Banitalebi-Dehkordi, Mo Chen, Yong Zhang arXiv ID 2203.15925 Category cs.RO: Robotics Cross-listed cs.AI, cs.LG, cs.MA Citations 2 Venue IEEE/RJS International Conference on Intelligent RObots and Systems Last Checked 4 months ago

Abstract

Cooperative multi-agent problems often require coordination between agents, which can be achieved through a centralized policy that considers the global state. Multi-agent policy gradient (MAPG) methods are commonly used to learn such policies, but they are often limited to problems with low-level action spaces. In complex problems with large state and action spaces, it is advantageous to extend MAPG methods to use higher-level actions, also known as options, to improve the policy search efficiency. However, multi-robot option executions are often asynchronous, that is, agents may select and complete their options at different time steps. This makes it difficult for MAPG methods to derive a centralized policy and evaluate its gradient, as centralized policy always select new options at the same time. In this work, we propose a novel, conditional reasoning approach to address this problem and demonstrate its effectiveness on representative option-based multi-agent cooperative tasks through empirical validation. Find code and videos at: \href{https://sites.google.com/view/mahrlsupp/}{https://sites.google.com/view/mahrlsupp/}

📄 View on arXiv 🌐 View on ar5iv 📑 PDF 🎉 Report Code Found

Community Contributions

Found the code? Know the venue? Think something is wrong? Let us know!

📜 Similar Papers

In the same crypt — Robotics

R.I.P. 👻 Ghosted

Past, Present, and Future of Simultaneous Localization And Mapping: Towards the Robust-Perception Age

Cesar Cadena, Luca Carlone, ... (+6 more)

cs.RO 🏛 IEEE TRO 📚 3.2K cites 10 years ago

R.I.P. 👻 Ghosted

AirSim: High-Fidelity Visual and Physical Simulation for Autonomous Vehicles

Shital Shah, Debadeepta Dey, ... (+2 more)

cs.RO 🏛 ICFSR 📚 2.3K cites 9 years ago

📚 📚 The Cartographer

A Survey of Motion Planning and Control Techniques for Self-driving Urban Vehicles

Brian Paden, Michal Cap, ... (+3 more)

cs.RO 🏛 IEEE TIV 📚 2.3K cites 10 years ago

📚 📚 The Cartographer

Unmanned Aerial Vehicles: A Survey on Civil Applications and Key Research Challenges

Hazim Shakhatreh, Ahmad Sawalmeh, ... (+7 more)

cs.RO 🏛 arXiv 📚 1.8K cites 8 years ago

📚 📚 The Cartographer

A Survey of Autonomous Driving: Common Practices and Emerging Technologies

Ekim Yurtsever, Jacob Lambert, ... (+2 more)

cs.RO 🏛 IEEE Access 📚 1.7K cites 7 years ago

R.I.P. 👻 Ghosted

Learning agile and dynamic motor skills for legged robots

Jemin Hwangbo, Joonho Lee, ... (+5 more)

cs.RO 🏛 Sci. Robot. 📚 1.6K cites 7 years ago

Died the same way — 👻 Ghosted

R.I.P. 👻 Ghosted

Federated Learning: Strategies for Improving Communication Efficiency

Jakub Konečný, H. Brendan McMahan, ... (+4 more)

cs.LG 🏛 arXiv 📚 5.2K cites 9 years ago

R.I.P. 👻 Ghosted

In-Datacenter Performance Analysis of a Tensor Processing Unit

Norman P. Jouppi, Cliff Young, ... (+73 more)

cs.AR 🏛 ISCA 📚 5.1K cites 9 years ago

R.I.P. 👻 Ghosted

Deep Convolutional Neural Networks for Computer-Aided Detection: CNN Architectures, Dataset Characteristics and Transfer Learning

Hoo-Chang Shin, Holger R. Roth, ... (+7 more)

cs.CV 🏛 IEEE TMI 📚 4.9K cites 10 years ago

R.I.P. 👻 Ghosted

Explanation in Artificial Intelligence: Insights from the Social Sciences

Tim Miller

cs.AI 🏛 AI 📚 4.9K cites 9 years ago