| Citation: | LIN Jiaqi①②, WANG Yong③, LIN Xin④, YAN Shi①②, XU Xin④, GU Jiangchun④, ZHANG Senbai⑤. A Multi-Agent Active-Inference Collaborative Decision-Making Method for Space-Air-Ground Integrated Networks[J]. Journal of Electronics & Information Technology. doi: 10.11999/JEIT260727 |
| [1] |
GUO Hongzhi, LI Jingyi, LIU Jiajia, et al. A survey on space-air-ground-sea integrated network security in 6G[J]. IEEE Communications Surveys & Tutorials, 2022, 24(1): 53–87. doi: 10.1109/COMST.2021.3131332.
|
| [2] |
王雪, 孟姝宇, 钱志鸿. 面向6G全域融合的智能接入关键技术综述[J]. 电子与信息学报, 2024, 46(5): 1613–1631. doi: 10.11999/JEIT231224.
WANG Xue, MENG Shuyu, and QIAN Zhihong. An overview of key technologies for intelligent access toward 6G full-domain convergence[J]. Journal of Electronics & Information Technology, 2024, 46(5): 1613–1631. doi: 10.11999/JEIT231224.
|
| [3] |
林佳琦, 钱琪杰, 钟旭东, 等. 面向6G的知识驱动“自智”网络架构[J]. 通信学报, 2025, 46(10): 40–62. doi: 10.11959/j.issn.1000-436x.2025159.
LIN Jiaqi, QIAN Qijie, ZHONG Xudong, et al. Knowledge-driven “self-intelligent” network architecture for 6G[J]. Journal on Communications, 2025, 46(10): 40–62. doi: 10.11959/j.issn.1000-436x.2025159.
|
| [4] |
SHI Rongye, YU Xin, WANG Yandong, et al. Symmetry-Informed MARL: A decentralized and cooperative UAV swarm control approach for communication coverage[J]. IEEE Transactions on Mobile Computing, 2025, 24(9): 8039–8056. doi: 10.1109/TMC.2025.3553285.
|
| [5] |
KIM G S, CHO Y, CHUNG J, et al. Quantum multi-agent reinforcement learning for cooperative mobile access in space-air-ground integrated networks[J]. IEEE Transactions on Mobile Computing, 2026, 25(1): 1200–1218. doi: 10.1109/TMC.2025.3599683.
|
| [6] |
AO Tianyong, LI Haoqiang, ZHANG Kaixin, et al. Heterogeneous UAVs trajectory optimization for post-disaster target search based on MARL with graph attention network[J]. IEEE Transactions on Vehicular Technology, 2026, 75(1): 1412–1426. doi: 10.1109/TVT.2025.3594534.
|
| [7] |
ZHANG Jianshu and WU Xiaofu. Cooperative jamming over DRL-based frequency hopping wireless communications: A one-leader multi-follower Stackelberg game approach[J]. IEEE Transactions on Information Forensics and Security, 2025, 20: 9220–9234. doi: 10.1109/TIFS.2025.3604229.
|
| [8] |
LIN Xin, LIU Aijun, HAN Chen, et al. Intelligent adaptive MIMO transmission for nonstationary communication environment: A deep reinforcement learning approach[J]. IEEE Transactions on Communications, 2025, 73(8): 5965–5979. doi: 10.1109/TCOMM.2025.3529263.
|
| [9] |
LI Peixuan, WANG Yichen, WANG Zhangnan, et al. Joint task offloading and resource allocation strategy for hybrid MEC-enabled LEO satellite networks: A hierarchical game approach[J]. IEEE Transactions on Communications, 2025, 73(5): 3150–3166. doi: 10.1109/TCOMM.2024.3478111.
|
| [10] |
CHEN Tianjiao, WANG Xiaoyun, HUA Meihui, et al. Incentive-based task offloading for digital twins in 6G native artificial intelligence networks: A learning approach[J]. Frontiers of Information Technology & Electronic Engineering, 2025, 26(2): 214–229. doi: 10.1631/FITEE.2400240.
|
| [11] |
WANG Kaidi, MA Yi, MASHHADI M B, et al. Convergence acceleration in wireless federated learning: A Stackelberg game approach[J]. IEEE Transactions on Vehicular Technology, 2025, 74(1): 714–729. doi: 10.1109/TVT.2024.3452933.
|
| [12] |
张鸿, 廖彧歆, 王汝言, 等. 面向密集场景的空天地网络资源分配算法[J]. 电子与信息学报, 2024, 46(5): 1968–1976. doi: 10.11999/JEIT231086.
ZHANG Hong, LIAO Yuxin, WANG Ruyan, et al. Resource allocation algorithm of space-air-ground integrated network for dense scenarios[J]. Journal of Electronics & Information Technology, 2024, 46(5): 1968–1976. doi: 10.11999/JEIT231086.
|
| [13] |
林艳, 夏开元, 张一晋. 基于生成对抗网络辅助多智能体强化学习的边缘计算网络联邦切片资源管理[J]. 电子与信息学报, 2025, 47(3): 666–677. doi: 10.11999/JEIT240773.
LIN Yan, XIA Kaiyuan, and ZHANG Yijin. Federated slicing resource management in edge computing networks based on GAN-assisted multi-agent reinforcement learning[J]. Journal of Electronics & Information Technology, 2025, 47(3): 666–677. doi: 10.11999/JEIT240773.
|
| [14] |
REZAZADEH F, CHERGUI H, and MANGUES-BAFALLUY J. Explanation-guided deep reinforcement learning for trustworthy 6G RAN slicing[C]. 2023 IEEE International Conference on Communications Workshops (ICC Workshops), Rome, Italy, 2023: 1026–1031. doi: 10.1109/ICCWorkshops57953.2023.10283684.
|
| [15] |
HONG J, VAN TU N, and HONG J W K. A comprehensive survey on LLM-based network management and operations[J]. International Journal of Network Management, 2025, 35(6): e70029. doi: 10.1002/nem.70029.
|
| [16] |
SUBRAMANIAN A, BHATTACHARJEE A, GUPTA S K, et al. A neuro-symbolic approach to multi-agent RL for interpretability and probabilistic decision making[C]. AAAI Spring Symposium on User-Aligned Assessment of Adaptive AI Systems, Stanford, USA, 2024. (查阅网上资料, 未找到本条文献信息, 请确认).
|
| [17] |
FRISTON K, FITZGERALD T, RIGOLI F, et al. Active inference: A process theory[J]. Neural Computation, 2017, 29(1): 1–49. doi: 10.1162/NECO_a_00912.
|
| [18] |
NEDIC A and OZDAGLAR A. Distributed subgradient methods for multi-agent optimization[J]. IEEE Transactions on Automatic Control, 2009, 54(1): 48–61. doi: 10.1109/TAC.2008.2009515.
|
| [19] |
FINN C, ABBEEL P, and LEVINE S. Model-agnostic meta-learning for fast adaptation of deep networks[C]. Proceedings of the 34th International Conference on Machine Learning, Sydney, Australia, 2017: 1126–1135.
|
| [20] |
HINTON G, VINYALS O, and DEAN J. Distilling the knowledge in a neural network[EB/OL]. https://arxiv.org/abs/1503.02531, 2015. doi: 10.48550/arXiv.1503.02531.
|