| 1 |
劳伦斯·弗里德曼. 威慑[M]. 莫盛凯, 译. 上海: 上海人民出版社, 2022.
|
| 2 |
Morgan P M. Deterrence now[M]. Cambridge: Cambridge University Press, 2003: 9.
|
| 3 |
Schelling T C. The strategy of conflict[M]. Cambridge: Harvard University Press, 1980.
|
| 4 |
Schelling T C. Arms and influence[M]. London: Yale University Press, 1966.
|
| 5 |
Wohlstetter A. Delicate balance of terror[J]. Foreign Affairs, 1959, 37 (2): 211.
doi: 10.1163/2468-1733_shafr_sim140110031
|
| 6 |
Harsanyi J C. Games with incomplete information played by “Bayesian” players, I-III, part I. the basic model[J]. Management Science, 1967, 14 (3): 159.
doi: 10.1007/978-94-017-2527-9_6
|
| 7 |
Harsanyi J C. Games with incomplete information played by “Bayesian” players, Part II. Bayesian equilibrium[J]. Management Science, 1968, 14 (5): 320.
doi: 10.1287/mnsc.14.5.320
|
| 8 |
Harsanyi J C. Games with incomplete information played by “Bayesian” players, Part III. The basic probability distribution of the Game[J]. Management Science, 1968, 14 (7): 486.
doi: 10.1287/mnsc.14.7.486
|
| 9 |
Smith J M. Evolution and the theory of games[M]. Cambridge: Cambridge University Press, 1982.
|
| 10 |
王永县, 向钢华. 基于博弈分析的军事威慑理论研究[J]. 清华大学学报(哲学社会科学版), 2005, 20 (5): 62.
|
| 11 |
向钢华, 王永县. 基于累积前景理论的有限理性威慑模型[J]. 系统工程, 2006, 24 (12): 107.
doi: 10.3969/j.issn.1001-4098.2006.12.021
|
| 12 |
Allison G T, Zelikow P. Essence of decision: explaining the cuban missile crisis[M]. 2nd ed. London: Pearson Education, Incorporated, 1999.
|
| 13 |
Karl S. Prospects of deterrence: deterrence theory, representation and evidence[J]. Defence and Peace Economics, 2024, 35 (2): 145.
doi: 10.1080/10242694.2022.2152956
|
| 14 |
Ma S, Tauman Y, Zeckhauser R. Deterrence games and the disruption of information[J]. International Journal of Game Theory, 2024, 53, 261.
doi: 10.1007/s00182-023-00870-3
|
| 15 |
Kaelbling L P, Littman M L, Cassandra A R. Planning and acting in partially observable stochastic domains[J]. Artificial Intelligence, 1998, 101 (1/2): 99.
doi: 10.1016/s0004-3702(98)00023-x
|
| 16 |
Clempner J B, Poznyak A. Optimization and games for controllable Markov chains: numerical methods with application to finance and engineering[M]. Gewerbestrasse: Springer Nature Switzerland AG, 2024: 47.
|
| 17 |
He K, Doshi P, Banerjee B. Modeling and reinforcement learning in partially observable many-agent systems[J]. Autonomous Agents and Multi-Agent Systems, 2024, 38(3): 12.
|
| 18 |
Schwartz J, Newbury R, Kulic D, et al. POSGGym: a library for decision-theoretic planning and learning in partially observable, multi-agent environments[J]. Autonomous Agents and Multi-Agent Systems, 2025, 39(7): 35.
|
| 19 |
Li J C, Cai M Y, Kan Z, et al. Model-free reinforcement learning for motion planning of autonomous agents with complex tasks in partially observable environments[J]. Autonomous Agents and Multi-Agent Systems, 2024, 38(4): 14.
|
| 20 |
Wang Z J, Wang B, Dou H B, et al. Windows deep transformer Q-networks: an extended variance reduction architecture for partially observable reinforcement learning[J]. Applied Intelligence, 2025, 55 (1): 35.
doi: 10.1007/s10489-024-05867-3
|
| 21 |
Helmeczi R K, Kavaklioglu C, Cevik M, et al. A multi-objective constrained partially observable Markov decision process model for breast cancer screening[J]. Operational Research, 2023, 23 (4): 30.
doi: 10.1007/s12351-023-00774-w
|
| 22 |
Kahneman D, Amos T. Prospect theory: an analysis of decision under risk[J]. Econometrica, 1979, 47 (2): 263.
doi: 10.21236/ada045771
|
| 23 |
王彪, 薛源, 陈萍萍, 等. 基于前景理论的多阶段多情景多部门应急决策的矩阵方法[J]. 控制与决策, 2025, 49 (2): 655.
doi: 10.13195/j.kzyjc.2023.1707
|