Article
Ansarinejad M., Gaweesh S. M., Ahmed M. M. (2025), Assessing the Efficacy of Pre-Trained Large Language Models in Analyzing Autonomous Vehicle Field Test Disengagements, Accident Analysis & Prevention, 220, Elsevier, 108178.
10.1016/j.aap.2025.108178Arai H. et al. (2025), CoVLA: Comprehensive Vision-Language-Action Dataset for Autonomous Driving, IEEE/CVF Winter Conference on Applications of Computer Vision, IEEE/CVF.
10.1109/WACV61041.2025.00195Bai S., Cai Y. X., Chen R. Z. et al. (2025), Qwen3-VL Technical Report, CoRR, abs/2511.21631, arXiv.
Cui C., Ma Y. T., Cao X. B., Ye W. Q., Zhou Y. F., Liang K. et al. (2024), A Survey on Multimodal Large Language Models for Autonomous Driving, 2024 IEEE/CVF Winter Conference on Applications of Computer Vision Workshops, IEEE, 958-979.
10.1109/WACVW60836.2024.00106Fu C. Y., Dai Y. L., Luo Y. H. et al. (2025), Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis, Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, IEEE/CVF, 24108-24118.
10.1109/CVPR52734.2025.02245Jeong E., Oh C. H., Lee S. S. (2017), Is Vehicle Automation Enough to Prevent Crashes? Role of Traffic Operations in Automated Driving Environments for Traffic Safety, Accident Analysis & Prevention, 104, Elsevier, 115-124.
10.1016/j.aap.2017.05.002Kim J. K., Rohrbach A., Darrell T., Canny J., Akata Z. (2018), Textual Explanations for Self-Driving Vehicles, Computer Vision - ECCV 2018, Springer, 577-593.
10.1007/978-3-030-01216-8_35Koopman P., Wagner M. (2016), Challenges in Autonomous Vehicle Testing and Validation, SAE International Journal of Transportation Safety, 4(1), SAE International, 15-24.
10.4271/2016-01-0128Laureshyn A., Svensson A., Hyden C. (2010), Evaluation of Traffic Safety, Based on Micro-level Behavioural Data: Theoretical Framework and First Implementation, Accident Analysis & Prevention, 42(6), Elsevier, 1637-1646.
10.1016/j.aap.2010.03.021Li K., He Y. N., Wang Y. X. et al. (2025), VideoChat: Chat-Centric Video Understanding, Science China Information Sciences, 68, Science China Press, 200102.
10.1007/s11432-024-4321-9Li K., Wang Y. X., He Y. N. et al. (2024), MVBench: A Comprehensive Multi-modal Video Understanding Benchmark, Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, IEEE/CVF, 22195-22206.
10.1109/CVPR52733.2024.02095Ma Y. T., Cui C., Cao X. B., Ye W. Q., Liu P., Lu J. J. et al. (2024), LaMPilot: An Open Benchmark Dataset for Autonomous Driving with Language Model Programs, Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, IEEE/CVF, 15141-15151.
10.1109/CVPR52733.2024.01434Marcu A. M., Chen L. Z., Huenermann J., Karnsund A., Hanotte B., Chidananda P. et al. (2024), LingoQA: Visual Question Answering for Autonomous Driving, Computer Vision - ECCV 2024, Springer, 252-269.
10.1007/978-3-031-72980-5_15Nie M. J., Peng R., Wang C., Cai X. Y., Han J. S., Xu H. Y., Zhang L. (2024), Reason2Drive: Towards Interpretable and Chain-Based Reasoning for Autonomous Driving, Computer Vision - ECCV 2024, Springer, 292-308.
10.1007/978-3-031-73347-5_17Papadoulis A., Quddus M., Imprialou M. (2019), Evaluating the Safety Impact of Connected and Autonomous Vehicles on Motorways, Accident Analysis & Prevention, 124, Elsevier, 12-22.
10.1016/j.aap.2018.12.019Qian T. B., Chen J. Y., Zhuo L., Jiao Y. H., Jiang Y. G. (2024), NuScenes-QA: A Multi-modal Visual Question Answering Benchmark for Autonomous Driving Scenario, Proceedings of the AAAI Conference on Artificial Intelligence, 38(5), AAAI Press, 4542-4550.
10.1609/aaai.v38i5.28253Sima C., Renz K., Chitta K., Chen L., Zhang H. J., Xie C. G. et al. (2024), DriveLM: Driving with Graph Visual Question Answering, Computer Vision - ECCV 2024, Springer, 256-274.
10.1007/978-3-031-72943-0_15Sun G. Z., Yang Y. D., Zhuang J. M. et al. (2025), video-SALMONN-o1: Reasoning-Enhanced Audio-Visual Large Language Model, CoRR, abs/2502.11775, arXiv.
Ulbrich S., Menzel T., Reschka A., Schuldt F., Maurer M. (2015), Defining and Substantiating the Terms Scene, Situation, and Scenario for Automated Driving, 2015 IEEE 18th International Conference on Intelligent Transportation Systems, IEEE.
10.1109/ITSC.2015.164Wang W. Y., Gao Z. W., Gu L. X. et al. (2025), InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency, CoRR, abs/2508.18265, arXiv.
Westhofen L., Neurohr C., Koopmann T., Butz M., Schuett B. U., Utesch F., Kramer B., Gutenkunst C., Boede E. (2023), Criticality Metrics for Automated Driving: A Review and Suitability Analysis of the State of the Art, Archives of Computational Methods in Engineering, Springer.
10.1007/s11831-022-09788-7Xie S. Z., Kong L. D., Dong Y., Sima C., Zhang W. J., Chen Q. A., Liu Z., Pan L. (2025), Are VLMs Ready for Autonomous Driving? An Empirical Study from the Reliability, Data, and Metric Perspectives, CoRR, abs/2501.04003, arXiv.
10.1109/ICCV51701.2025.00621Xue L., Shu M. L., Awadalla A. et al. (2024), xGen-MM (BLIP-3): A Family of Open Large Multimodal Models, CoRR, abs/2408.08872, arXiv.
Yan Z. A., Li X. H., He Y. N. et al. (2025), VideoChat-R1.5: Visual Test-Time Scaling to Reinforce Multimodal Reasoning by Iterative Perception, Advances in Neural Information Processing Systems, Curran Associates.
10.52202/085713-3976Z.ai (2026a), GLM-4.6V-Flash Model Card, Hugging Face, https://huggingface.co/zai-org/GLM-4.6V-Flash.
Z.ai (2026b), GLM-V: GLM-4.6V/4.5V/4.1V-Thinking, GitHub Repository, https://github.com/zai-org/GLM-V.
- Publisher :Korean Society of Transportation
- Publisher(Ko) :대한교통학회
- Journal Title :Journal of Korean Society of Transportation
- Journal Title(Ko) :대한교통학회지
- Volume : 44
- No :4
- Pages :722-736
- Received Date : 2026-05-19
- Revised Date : 2026-06-01
- Accepted Date : 2026-06-08
- DOI :https://doi.org/10.7470/jkst.2026.44.4.722


Journal of Korean Society of Transportation







