MABL: Bi-Level Latent-Variable World Model for Sample-Efficient Multi-Agent Reinforcement Learning | allainews.com

Feb. 15, 2024, 5:43 a.m. | Aravind Venugopal, Stephanie Milani, Fei Fang, Balaraman Ravindran

cs.LG updates on arXiv.org arxiv.org

arXiv:2304.06011v2 Announce Type: replace
Abstract: Multi-agent reinforcement learning (MARL) methods often suffer from high sample complexity, limiting their use in real-world problems where data is sparse or expensive to collect. Although latent-variable world models have been employed to address this issue by generating abundant synthetic data for MARL training, most of these models cannot encode vital global information available during training into their latent states, which hampers learning efficiency. The few exceptions that incorporate global information assume centralized execution of …

abstract agent arxiv complexity cs.lg cs.ma data issue multi-agent reinforcement reinforcement learning sample synthetic synthetic data training type world world models

More from arxiv.org / cs.LG updates on arXiv.org

Red-Teaming for Generative AI: Silver Bullet or Security Theater? 2 days, 19 hours ago | arxiv.org

abstract arxiv concerns cs.cy +15

Efficient Data-Driven MPC for Demand Response of Commercial Buildings 2 days, 19 hours ago | arxiv.org

abstract arxiv buildings commercial +20

BrepGen: A B-rep Generative Diffusion Model with Structured Latent Geometry 2 days, 19 hours ago | arxiv.org

arxiv cs.cv cs.lg diffusion +5

Data-Driven Physics-Informed Neural Networks: A Digital Twin Perspective 2 days, 19 hours ago | arxiv.org

abstract arxiv automated construction +26

Testing the Segment Anything Model on radiology data 2 days, 19 hours ago | arxiv.org

abstract applications arxiv become +20

Robust Point Matching with Distance Profiles 2 days, 19 hours ago | arxiv.org

abstract analyze arxiv cs.lg +13

Cell Maps Representation For Lung Adenocarcinoma Growth Patterns Classification In Whole Slide Images 2 days, 19 hours ago | arxiv.org

abstract arxiv behavior classification +18

Improved Baselines with Visual Instruction Tuning 2 days, 19 hours ago | arxiv.org

abstract academic arxiv clip +25

Calorimeter shower superresolution 2 days, 19 hours ago | arxiv.org

abstract arxiv challenge computational +16

Software Engineer for AI Training Data (School Specific)

@ G2i Inc | Remote

View on ai-jobs.net

Software Engineer for AI Training Data (Python)

@ G2i Inc | Remote

View on ai-jobs.net

Software Engineer for AI Training Data (Tier 2)

@ G2i Inc | Remote

View on ai-jobs.net

Data Engineer

@ Lemon.io | Remote: Europe, LATAM, Canada, UK, Asia, Oceania

View on ai-jobs.net

Artificial Intelligence – Bioinformatic Expert

@ University of Texas Medical Branch | Galveston, TX

View on ai-jobs.net

Lead Developer (AI)

@ Cere Network | San Francisco, US

View on ai-jobs.net