| 1 |
赵伟, 叶军, 王邠. 基于人工智能的智能化指挥决策和控制[J]. 信息安全与通信保密, 2022 (2): 2.
|
| 2 |
耿欣. 面向智能杀伤链中的美军态势感知体系研究[J]. 情报杂志, 2025, 44 (3): 12.
doi: 10.3969/j.issn.1002-1965.2025.03.002
|
| 3 |
王晓丹, 向前, 李睿, 等. 深度学习研究及军事应用综述[J]. 空军工程大学学报(自然科学版), 2022, 23 (1): 1.
|
| 4 |
张悦, 孙岱. 指挥决策中应急情报的发展思路[J]. 电信快报, 2023 (6): 21.
doi: 10.3969/j.issn.1006-1339.2023.06.006
|
| 5 |
Wang X, Wang S, Liang X X, et al. Deep reinforcement learning: a survey[J]. IEEE Trans. on Neural Networks and Learning Systems, 2022, 35 (4): 5064.
|
| 6 |
张梦钰, 豆亚杰, 陈子夷, 等. 深度强化学习及其在军事领域中的应用综述[J]. 系统工程与电子技术, 2024, 46 (4): 1297.
doi: 10.12305/j.issn.1001-506X.2024.04.18
|
| 7 |
Zhang K Q. Foundations of multi-agent learning in dynamic environments: where reinforcement learning meets strategic decision-making[C]//2025 39th AAAI Conference on Artificial Intelligence, 2025: 28734.
|
| 8 |
Cao J B, Dou J F, Liu J L, et al. Multi-agent deep reinforcement learning framework strategized by unmanned aerial vehicles for multi-vessel full communication connection[J]. Remote Sensing, 2023, 15 (16): 4059.
doi: 10.3390/rs15164059
|
| 9 |
Wei X L, Yang L F, Cao G, et al. Recurrent MADDPG for object detection and assignment in combat tasks[J]. IEEE Access, 2020, 8, 163334.
doi: 10.1109/ACCESS.2020.3022638
|
| 10 |
Yoon S, Cho J H, Kim D S, et al. DESOLATER: deep reinforcement learning-based resource allocation and moving target defense deployment framework[J]. IEEE Access, 2021, 9, 70700.
doi: 10.1109/ACCESS.2021.3076599
|
| 11 |
Li R Z, Yuan H, Ren B B, et al. Optimal unmanned combat system-of-systems reconstruction strategy with heterogeneous cost via deep reinforcement learning[J]. Mathematics, 2024, 12 (10): 1476.
doi: 10.3390/math12101476
|
| 12 |
Vasankari L, Saastamoinen K. Strategizing the shallows: leveraging multi-agent reinforcement learning for enhanced tactical decision-making in littoral naval warfare[C]//IFIP International Conference on Artificial Intelligence Applications and Innovations, 2024: 129.
|
| 13 |
吴冯国, 陶伟, 李辉, 等. 基于深度强化学习算法的无人机智能规避决策[J]. 系统工程与电子技术, 2023, 45 (6): 1702.
|
| 14 |
林志康, 施龙飞, 刘甲磊, 等. 基于深度Q学习的组网雷达闪烁探测调度方法[J]. 系统工程与电子技术, 2025, 47 (5): 1443.
|
| 15 |
秦湖程, 黄炎焱, 陈天德, 等. 基于PPO算法的集群多目标火力规划方法[J]. 系统工程与电子技术, 2024, 46 (11): 3764.
doi: 10.12305/j.issn.1001-506X.2024.11.18
|
| 16 |
张庭瑜, 曾颖, 李楠, 等. 基于深度强化学习的航天器功率-信号复合网络优化算法[J]. 系统工程与电子技术, 2024, 46 (9): 3060.
|
| 17 |
夏雨奇, 黄炎焱, 陈恰. 基于深度Q网络的无人车侦察路径规划[J]. 系统工程与电子技术, 2024, 46 (9): 3070.
|
| 18 |
张杰, 王刚, 宋亚飞, 等. 基于自适应SGD-多智能体的防空资源部署优化[J]. 系统工程与电子技术, 2019, 41 (7): 1536.
|
| 19 |
王子怡, 傅雄军, 董健, 等. 基于分层多智能体强化学习的雷达协同抗干扰策略优化[J]. 系统工程与电子技术, 2025, 47 (4): 1108.
doi: 10.12305/j.issn.1001-506X.2025.04.07
|
| 20 |
马悦, 吴琳, 许霄. 基于多智能体强化学习的协同目标分配[J]. 系统工程与电子技术, 2023, 45 (9): 2793.
doi: 10.12305/j.issn.1001-506X.2023.09.18
|
| 21 |
赵芷若, 曹雷, 陈希亮, 等. 基于多智能体博弈强化学习的无人机智能攻击策略生成模型[J]. 系统工程与电子技术, 2023, 45 (10): 3165.
|
| 22 |
Zhang J D, Yang Q M, Shi G Q, et al. UAV cooperative air combat maneuver decision based on multi-agent reinforcement learning[J]. Journal of Systems Engineering and Electronics, 2021, 32 (6): 1421.
doi: 10.23919/jsee.2021.000121
|
| 23 |
Ladosz P, Weng L, Kim M, et al. Exploration in deep reinforcement learning: a survey[J]. Information Fusion, 2022, 85, 1.
|
| 24 |
Sumiea E H, Abdulkadir S J, Alhussian H S, et al. Deep deterministic policy gradient algorithm: a systematic review[J]. Heliyon, 2024, 10 (6): e27942.
|
| 25 |
Jia Y W, Zhou X Y. Policy gradient and actor-critic learning in continuous time and space: theory and algorithms[J]. Journal of Machine Learning Research, 2022, 23, 1.
|
| 26 |
Li M X, Wang Q, Xu Y J. GTDE: grouped training with decentralized execution for multi-agent actor-critic[C]//2022 39th AAAI Conference on Artificial Intelligence, 2025: 18368.
|
| 27 |
Ma Y, Liu Y G, Zhao L, et al. A review on cooperative control problems of multi-agent systems[C]//41st Chinese Control Conference , 2022: 4831.
|
| 28 |
Pandya R, Zhao M, Liu C, et al. Multi-agent strategy explanations for human-robot collaboration[C]//IEEE International Conference on Robotics and Automation, 2024: 17351.
|
| 29 |
Hu T M, Luo B, Yang C H, et al. MO-MIX: multi-objective multi-agent cooperative decision-making with deep reinforcement learning[J]. IEEE Trans. on Pattern Analysis and Machine Intelligence, 2023, 45 (10): 12098.
doi: 10.1109/TPAMI.2023.3283537
|
| 30 |
Oroojlooy A, Hajinezhad D. A review of cooperative multi-agent deep reinforcement learning[J]. Applied Intelligence, 2023, 53 (11): 13677.
doi: 10.1007/s10489-022-04105-y
|
| 31 |
Lee C E, Baek J, Do S W, et al. Multi-agent based collaborative agent architecture for battlefield situation awareness[C]//IEEE International Conference on Big Data and Smart Computing, 2024: 395.
|
| 32 |
Suilen M, Simao T D, Parker D, et al. Robust anytime learning of Markov decision processes[J]. Advances in Neural Information Processing Systems, 2022, 35, 28790.
doi: 10.52202/068431-2087
|
| 33 |
Kurniawati H. Partially observable Markov decision processes and robotics[J]. Annual Review of Control, Robotics, and Autonomous Systems, 2022, 5 (1): 253.
doi: 10.1146/annurev-control-042920-092451
|
| 34 |
Tan C S, Van Bossuyt D L, Hale B. System analysis of counter-unmanned aerial systems kill chain in an operational environment[J]. Systems, 2021, 9 (4): 79.
doi: 10.3390/systems9040079
|
| 35 |
秦长江, 吴克宇, 成清, 等. 基于杀伤网贡献率的动态体系节点重要度评估[J]. 系统工程与电子技术, 2023, 45 (6): 1732.
doi: 10.12305/j.issn.1001-506X.2023.06.17
|
| 36 |
王耀祖, 尚柏林, 宋笔锋, 等. 基于杀伤链的作战体系网络关键节点识别方法[J]. 系统工程与电子技术, 2023, 45 (3): 736.
|
| 37 |
韩明磊, 马晶, 周泽宇, 等. 基于Agent建模的海战场杀伤链评估系统研究[J]. 计算机仿真, 2022, 39 (3): 11.
|
| 38 |
Huang W H, Li K, Shao K, et al. Multi-agent Q-learning with sub-team coordination[J]. Advances in Neural Information Processing Systems, 2022, 35, 29427.
|
| 39 |
Zhang J, Zhang Y, Zhang X S, et al. Intrinsic action tendency consistency for cooperative multi-agent reinforcement learning[C]//AAAI Conference on Artificial Intelligence, 2024: 17600.
|
| 40 |
Wu Z F, Yu C, Ye D H, et al. Coordinated proximal policy optimization[J]. Advances in Neural Information Processing Systems, 2021, 34, 26437.
|