[1] CHAKRABORTI T, KULKARNI A, SREEDHARAN S, et al. Explicability? Legibility? Predictability? Transparency? Privacy? Security? The emerging landscape of interpretable agent behavior[C] //Proceedings of the 29th International Conference on Automated Planning and Scheduling (ICAPS 2019). 2019, 29(1): 86-96.
[2] 丁世飞,杜威,张健,等.多智能体深度强化学习研究进展[J].计算机学报,2024,47(7):1547-1567.
DING S F, DU W, ZHANG J, et al. Research Progress of Multi-Agent Deep Reinforcement Learning[J]. Chinese Journal of Computers, 2024, 47(7): 1547-1567. (in Chinese)
[3] 毕千, 钱程, 张可, 等. 基于深度强化学习的多智能体角度跟踪方法设计[J]. 计算机工程, 2024, 50(11): 10-17.
BI Q, QIAN C, ZHANG K, et al. Design of multi-agent angle tracking method based on deep reinforcement learning[J]. Computer Engineering, 2024, 50(11): 10-17. (in Chinese)
[4] SCHMIDT-WOLF M, BECKER T, OLIVA D, et al. Investigating non-verbal cues in cluttered environments: insights into legible motion from interpersonal interaction[C]//Proceedings of the 2024 33rd IEEE International Conference on Robot and Human Interactive Communication (RO-MAN). 2024: 1250-1257.
[5] BAKER C L, SAXE R, TENENBAUM J B. Action understanding as inverse planning[J]. Cognition, 2009, 113(3): 329-349.
[6] ZHANG Z, ZENG Y F, JIANG W H, et al. Intention recognition for multiple agents[J]. Information Sciences, 2023, 628: 360-376.
[7] 李雪松, 张锲石, 宋呈群, 等. 自动驾驶场景下的轨迹预测技术综述[J]. 计算机工程, 2023, 49(5): 1-11.
LI X S, ZHANG Q S, SONG C Q, et al. Review of Trajectory Prediction Technology in Autonomous Driving Scenes[J]. Computer Engineering, 2023, 49(5): 1-11. (in Chinese)
[8] MIURA S, COHEN A L, ZILBERSTEIN S. Maximizing legibility in stochastic environments[C]//Proceedings of the 2021 30th IEEE International Conference on Robot and Human Interactive Communication (RO-MAN). Piscataway: IEEE, 2021: 1053-1059.
[9] FARIA M, MELO F S, PAIVA A. "Guess what I'm doing": extending legibility to sequential decision tasks[J]. Artificial Intelligence, 2024, 330: 104107.
[10] 张红强,石佳航,吴亮红,等. 改进MADDPG算法的非凸环境下多智能体自组织协同围捕[J]. 计算机科学与探索, 2024, 18 (08): 2080-2090.
ZHANG H Q, SHI J H, WU L H, et al. Multi-agent Self-organizing Cooperative Hunting in Non-convex Environment with Improved MADDPG Algorithm[J]. Journal of Frontiers of Computer Science and Technology, 2024, 18(08): 2080-2090. (in Chinese)
[11] MIURA S, ZILBERSTEIN S. A unifying framework for observer-aware planning and its complexity[C] //Proceedings of the Thirty-Seventh Conference on Uncertainty in Artificial Intelligence (UAI 2021). PMLR, 2021: 161: 610-620.
[12] LIU Y, PAN Y, ZENG Y, et al. Active legibility in multiagent reinforcement learning[J]. Artificial Intelligence, 2025, 346: 104357.
[13] BERNARDINI S, FAGNANI F, NEACSU A, et al. Optimizing pathfinding for goal legibility and recognition in cooperative partially observable environments[J]. Artificial Intelligence, 2024, 333: 104148.
[14] ZENG B X, ZENG Y F, PAN Y H. Inverse reinforcement learning for legibility automation in intelligent agents[C]//2024 IEEE Conference on Artificial Intelligence (CAI). Singapore: IEEE, 2024: 741-746.
[15] LIU Y Y, ZENG Y F, MA B Y, et al. Improvement and evaluation of the policy legibility in reinforcement learning[C]//Proceedings of the 22nd International Conference on Autonomous Agents and Multiagent Systems (AAMAS). 2023: 3044-3046.
[16] YANG T, SHI X H, ZENG Q H, et al. Optimization methods in fully cooperative scenarios: a review of multi-agent reinforcement learning[J]. Frontiers of Information Technology & Electronic Engineering, 2025, 26(4): 479-509.
[17] DRAGAN A D, BAUMAN S, FORLIZZI J, et al. Effects of robot motion on human-robot collaboration[C] //Proceedings of the 2015 ACM/IEEE International Conference on Human-Robot Interaction (HRI'15). USA: ACM, 2015: 51-58.
[18] 沈思彤,王耀吾,谢在鹏,唐斌。基于角色学习的多智能体强化学习方法 [J]. 计算机工程,2025, 51 (6): 102-115.
SHEN S T, WANG Y W, XIE Z P, TANG B. Role Learning-based Multi-Agent Reinforcement Learning Methods [J]. Computer Engineering, 2025, 51 (6): 102-115. (in Chinese)
[19] DRAGAN A D, LEE K C T, SRINIVASA S. Legibility and predictability of robot motion[C]//2013 8th ACM/IEEE International Conference on Human-Robot Interaction (HRI). Piscataway: IEEE, 2013: 301-308.
[20] DRAGAN A, SRINIVASA S. Generating legible motion[C]//Proceedings of Robotics: Science and Systems (RSS '13). USA: MIT Press, 2013: 7732.
[21] MARINHO Z, DRAGAN A, BYRAVAN A, et al. Functional gradient motion planning in reproducing kernel Hilbert spaces[EB/OL].(2016-01-14)[2026-03-03] https://arxiv.org/abs/1601.03648.
[22] BRONARS M, XU D. Legible robot motion from conditional generative models[C]//Proceedings of the 40th International Conference on Machine Learning (ICML 2023). USA: PMLR, 2023.
[23] 钱诚泽,毛剑琳,李睿褀,等. 基于冲突代Bayesian权重的改进PBS多智能体路径规划算法[J]. 电子学报, 2025: 1-14.
QIAN C, MAO J, LI R, et al. Improved PBS Multi-Agent Path Planning Algorithm Based on Conflict Cost Bayesian Weighting [J]. Acta Electronica Sinica, 2025:1-14. (in Chinese)
[24] LI W, LIU W Y, SHAO S T, et al. Attention-based intrinsic reward mixing network for credit assignment in multi-agent reinforcement learning[J]. IEEE Transactions on Games, 2024, 16(2): 270-281.
[25] WALLKOTTER S, CHETOUANI M, CASTELLANO G. A new approach to evaluating legibility: comparing legibility frameworks using framework-independent robot motion trajectories [EB/OL]. (2022-01-15) [2026-03-03]. https://arxiv.org/abs/2201.05765v1.
[26] TAYLOR A V, MAMANTOV E, ADMONI H. Observer-aware legibility for social navigation[C]//2022 31st IEEE International Conference on Robot and Human Interactive Communication (RO-MAN). Italy: IEEE, 2022: 1115-1122.
[27] SCHMIDT-WOLF M, BECKER T J, OLIVA D, et al. Through the clutter: exploring the impact of complex environments on the legibility of robot motion[C]//2025 IEEE International Conference on Robotics and Automation (ICRA). USA: IEEE, 2025: 13553-13559.
[28] DAS A, GERVET T, ROMOFF J, et al. Tarmac: targeted multi-agent communication [C]//Proceedings of the International Conference on Machine Learning (ICML 2019).: PMLR, 2019: 1538-1546.
[29] 林海波,王浩,张毅. 改进高斯核函数的人体姿态分析与识别 [J]. 智能系统学报, 2015, 10 (3): 436-441.
LIN H B, WANG H, ZHANG Y. Human postures recognition based on the improved Gauss kernel function[J]. CAAI Transactions on Intelligent Systems, 2015, 10(3): 436-441. (in Chinese)
[30] PERSIANI M, HELLSTRÖM T. Policy regularization for legible behavior[J]. Neural Computing and Applications, 2023, 35(23): 16781-16790.
|