All Issue

2026 Vol.44, Issue 4 Preview Page

Article

31 August 2026. pp. 722-736
Abstract
References
1

Ansarinejad M., Gaweesh S. M., Ahmed M. M. (2025), Assessing the Efficacy of Pre-Trained Large Language Models in Analyzing Autonomous Vehicle Field Test Disengagements, Accident Analysis & Prevention, 220, Elsevier, 108178.

10.1016/j.aap.2025.108178
2

Arai H. et al. (2025), CoVLA: Comprehensive Vision-Language-Action Dataset for Autonomous Driving, IEEE/CVF Winter Conference on Applications of Computer Vision, IEEE/CVF.

10.1109/WACV61041.2025.00195
3

Bai S., Cai Y. X., Chen R. Z. et al. (2025), Qwen3-VL Technical Report, CoRR, abs/2511.21631, arXiv.

4

Cui C., Ma Y. T., Cao X. B., Ye W. Q., Zhou Y. F., Liang K. et al. (2024), A Survey on Multimodal Large Language Models for Autonomous Driving, 2024 IEEE/CVF Winter Conference on Applications of Computer Vision Workshops, IEEE, 958-979.

10.1109/WACVW60836.2024.00106
5

Fu C. Y., Dai Y. L., Luo Y. H. et al. (2025), Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis, Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, IEEE/CVF, 24108-24118.

10.1109/CVPR52734.2025.02245
6

Jeong E., Oh C. H., Lee S. S. (2017), Is Vehicle Automation Enough to Prevent Crashes? Role of Traffic Operations in Automated Driving Environments for Traffic Safety, Accident Analysis & Prevention, 104, Elsevier, 115-124.

10.1016/j.aap.2017.05.002
7

Kim J. K., Rohrbach A., Darrell T., Canny J., Akata Z. (2018), Textual Explanations for Self-Driving Vehicles, Computer Vision - ECCV 2018, Springer, 577-593.

10.1007/978-3-030-01216-8_35
8

Koopman P., Wagner M. (2016), Challenges in Autonomous Vehicle Testing and Validation, SAE International Journal of Transportation Safety, 4(1), SAE International, 15-24.

10.4271/2016-01-0128
9

Laureshyn A., Svensson A., Hyden C. (2010), Evaluation of Traffic Safety, Based on Micro-level Behavioural Data: Theoretical Framework and First Implementation, Accident Analysis & Prevention, 42(6), Elsevier, 1637-1646.

10.1016/j.aap.2010.03.021
10

Li K., He Y. N., Wang Y. X. et al. (2025), VideoChat: Chat-Centric Video Understanding, Science China Information Sciences, 68, Science China Press, 200102.

10.1007/s11432-024-4321-9
11

Li K., Wang Y. X., He Y. N. et al. (2024), MVBench: A Comprehensive Multi-modal Video Understanding Benchmark, Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, IEEE/CVF, 22195-22206.

10.1109/CVPR52733.2024.02095
12

Ma Y. T., Cui C., Cao X. B., Ye W. Q., Liu P., Lu J. J. et al. (2024), LaMPilot: An Open Benchmark Dataset for Autonomous Driving with Language Model Programs, Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, IEEE/CVF, 15141-15151.

10.1109/CVPR52733.2024.01434
13

Marcu A. M., Chen L. Z., Huenermann J., Karnsund A., Hanotte B., Chidananda P. et al. (2024), LingoQA: Visual Question Answering for Autonomous Driving, Computer Vision - ECCV 2024, Springer, 252-269.

10.1007/978-3-031-72980-5_15
14

Nie M. J., Peng R., Wang C., Cai X. Y., Han J. S., Xu H. Y., Zhang L. (2024), Reason2Drive: Towards Interpretable and Chain-Based Reasoning for Autonomous Driving, Computer Vision - ECCV 2024, Springer, 292-308.

10.1007/978-3-031-73347-5_17
15

Papadoulis A., Quddus M., Imprialou M. (2019), Evaluating the Safety Impact of Connected and Autonomous Vehicles on Motorways, Accident Analysis & Prevention, 124, Elsevier, 12-22.

10.1016/j.aap.2018.12.019
16

Qian T. B., Chen J. Y., Zhuo L., Jiao Y. H., Jiang Y. G. (2024), NuScenes-QA: A Multi-modal Visual Question Answering Benchmark for Autonomous Driving Scenario, Proceedings of the AAAI Conference on Artificial Intelligence, 38(5), AAAI Press, 4542-4550.

10.1609/aaai.v38i5.28253
17

Sima C., Renz K., Chitta K., Chen L., Zhang H. J., Xie C. G. et al. (2024), DriveLM: Driving with Graph Visual Question Answering, Computer Vision - ECCV 2024, Springer, 256-274.

10.1007/978-3-031-72943-0_15
18

Sun G. Z., Yang Y. D., Zhuang J. M. et al. (2025), video-SALMONN-o1: Reasoning-Enhanced Audio-Visual Large Language Model, CoRR, abs/2502.11775, arXiv.

19

Ulbrich S., Menzel T., Reschka A., Schuldt F., Maurer M. (2015), Defining and Substantiating the Terms Scene, Situation, and Scenario for Automated Driving, 2015 IEEE 18th International Conference on Intelligent Transportation Systems, IEEE.

10.1109/ITSC.2015.164
20

Wang W. Y., Gao Z. W., Gu L. X. et al. (2025), InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency, CoRR, abs/2508.18265, arXiv.

21

Westhofen L., Neurohr C., Koopmann T., Butz M., Schuett B. U., Utesch F., Kramer B., Gutenkunst C., Boede E. (2023), Criticality Metrics for Automated Driving: A Review and Suitability Analysis of the State of the Art, Archives of Computational Methods in Engineering, Springer.

10.1007/s11831-022-09788-7
22

Xie S. Z., Kong L. D., Dong Y., Sima C., Zhang W. J., Chen Q. A., Liu Z., Pan L. (2025), Are VLMs Ready for Autonomous Driving? An Empirical Study from the Reliability, Data, and Metric Perspectives, CoRR, abs/2501.04003, arXiv.

10.1109/ICCV51701.2025.00621
23

Xue L., Shu M. L., Awadalla A. et al. (2024), xGen-MM (BLIP-3): A Family of Open Large Multimodal Models, CoRR, abs/2408.08872, arXiv.

24

Yan Z. A., Li X. H., He Y. N. et al. (2025), VideoChat-R1.5: Visual Test-Time Scaling to Reinforce Multimodal Reasoning by Iterative Perception, Advances in Neural Information Processing Systems, Curran Associates.

10.52202/085713-3976
25

Z.ai (2026a), GLM-4.6V-Flash Model Card, Hugging Face, https://huggingface.co/zai-org/GLM-4.6V-Flash.

26

Z.ai (2026b), GLM-V: GLM-4.6V/4.5V/4.1V-Thinking, GitHub Repository, https://github.com/zai-org/GLM-V.

27

Zhang B. Q., Li K. H., Cheng Z. S. et al. (2025a), VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding, CoRR, abs/2501.13106, arXiv.

28

Zhang R. X., Wang B. C., Zhang J. X., Bian Z. L., Feng C., Ozbay K. (2025b), When Language and Vision Meet Road Safety: Leveraging Multimodal Large Language Models for Video-Based Traffic Accident Analysis, Accident Analysis & Prevention, 219, Elsevier, 108077.

10.1016/j.aap.2025.108077
Information
  • Publisher :Korean Society of Transportation
  • Publisher(Ko) :대한교통학회
  • Journal Title :Journal of Korean Society of Transportation
  • Journal Title(Ko) :대한교통학회지
  • Volume : 44
  • No :4
  • Pages :722-736
  • Received Date : 2026-05-19
  • Revised Date : 2026-06-01
  • Accepted Date : 2026-06-08