SyncAI.news, a Varaisys broadcasting
ICMAPE: In-Context Multiagent Pure Exploration
XH

Xinyi Hu, Alessio Russo, Aldo Pacchiano

· 1 min read

ResearcharXiv cs.LG

ICMAPE: In-Context Multiagent Pure Exploration

arXiv:2609.33986v1 Announce Type: new Abstract: In some multi-agent systems, the quantity to be optimized is not an externally specified reward but the information acquired about unknown properties of the environment as done in active sequential hypothesis testing (ASHT) problems. However, the ASHT literature tends to focus on finite single-agent problems with well-specified models, while there is currently a gap for practical multi-agent methods that can perform active sequential testing. We fill this gap with ICMAPE, a Bayesian learning-based framework for decentralized multi-agent pure-exploration driven by inference objectives. ICMAPE converts the fixed-confidence identification objective into a reward derived from inference confidence, so that standard reinforcement learning machinery can be applied to decentralized pure exploration. It jointly learns a centralized neural inference network that estimates a posterior distribution over hypotheses from global trajectory data, and decentralized policies that select actions from local observation histories and learn when to stop collecting data once the target confidence is reached. On two synthetic benchmarks and a Maryland nitrate concentration monitoring task based on real-world data, ICMAPE-TD3 achieves target accuracy with fewer exploration steps.

Original source

This story was published by arXiv cs.LG and written by Xinyi Hu, Alessio Russo, Aldo Pacchiano. SyncAI.news shows a preview; the complete article is on the publisher's site.

Read the full story on arxiv.org

Similar News