4.8 Article

Few-Shot Multi-Agent Perception With Ranking-Based Feature Learning

出版社

IEEE COMPUTER SOC
DOI: 10.1109/TPAMI.2023.3285755

关键词

Few-shot learning; image and audio classification; multi-agent perception; optimal transport; semantic segmentation

向作者/读者索取更多资源

In this article, a metric-based multi-agent few-shot learning framework is proposed, which enables agents to accurately and efficiently perceive the environment under limited communication and computation conditions through an efficient communication mechanism, asymmetric attention mechanism, and metric-learning module. Additionally, a specially designed ranking-based feature learning module is utilized to improve accuracy by maximizing inter-class distance and minimizing intra-class distance.
In this article, we focus on performing few-shot learning (FSL) under multi-agent scenarios in which participating agents only have scarce labeled data and need to collaborate to predict labels of query observations. We aim at designing a coordination and learning framework in which multiple agents, such as drones and robots, can collectively perceive the environment accurately and efficiently under limited communication and computation conditions. We propose a metric-based multi-agent FSL framework which has three main components: an efficient communication mechanism that propagates compact and fine-grained query feature maps from query agents to support agents; an asymmetric attention mechanism that computes region-level attention weights between query and support feature maps; and a metric-learning module which calculates the image-level relevance between query and support data fast and accurately. Furthermore, we propose a specially designed ranking-based feature learning module, which can fully utilize the order information of training data by maximizing the inter-class distance, while minimizing the intra-class distance explicitly. We perform extensive numerical studies and demonstrate that our approach can achieve significantly improved accuracy in visual and acoustic perception tasks such as face identification, semantic segmentation, and sound genre recognition, consistently outperforming the state-of-the-art baselines by 5%-20%.

作者

我是这篇论文的作者
点击您的名字以认领此论文并将其添加到您的个人资料中。

评论

主要评分

4.8
评分不足

次要评分

新颖性
-
重要性
-
科学严谨性
-
评价这篇论文

推荐

暂无数据
暂无数据