Author
Listed:
- Zhiheng Chen
(Northwest Institute of Mechanical and Electrical Engineering, China)
- Liang Ma
(Southeast University, China)
- Xiaoyu Cui
(Northwest Institute of Mechanical and Electrical Engineering, China)
- Zhen Wang
(Northwest Institute of Mechanical and Electrical Engineering, China)
- Jialin Yao
(Northwest Institute of Mechanical and Electrical Engineering, China)
- Jiawei Hu
(Northwest Institute of Mechanical and Electrical Engineering, China)
Abstract
In complex traffic environments, conventional vehicle detection methods often show limited precision and robustness when facing distant, small-scale, and occluded vehicles. To address these issues, this research proposes a Multiscale, Multi knowledge, Attention enhanced You Only Look Once (YOLO) model. It is a knowledge-driven multiscale vehicle detection framework for intelligent transportation systems, built on the YOLO version 8 nano model. The framework introduces three modules. The first is a cross-stage partial mixed aggregation network module. This module enhances backbone representation through dynamic multiscale aggregation. The second is a multiscale downsampling module that combines dilated convolution and parallel pooling to preserve vehicle cues during downsampling. The third module is an aspect-ratio perception with cross-attention module that injects aspect-ratio-aware knowledge with cross-attention to adapt to vehicle shape variation and suppress background interference. Experiments on the Vehicle dataset on Kaggle.com and the Berkeley DeepDrive 100K dataset showed that the Multiscale, Multi knowledge, Attention enhanced-YOLO model outperformed the lightweight YOLO baseline models. On the Berkeley DeepDrive 100,000 dataset, it achieved a 48.17% mean average precision @0.5 and 26.70% mean average precision @0.5:0.95, improving YOLO version 8 nano by 3.34% and 2.19%, respectively. Visualization results further confirmed its effectiveness under occlusion and adverse-weather conditions while maintaining real-time efficiency.
Suggested Citation
Zhiheng Chen & Liang Ma & Xiaoyu Cui & Zhen Wang & Jialin Yao & Jiawei Hu, 2026.
"Knowledge-Driven Multi-Scale Vehicle Detection Framework for Intelligent Transportation Systems,"
International Journal on Semantic Web and Information Systems (IJSWIS), IGI Global Scientific Publishing, vol. 22(1), pages 1-35, January.
Handle:
RePEc:igg:jswis0:v:22:y:2026:i:1:p:1-35
Download full text from publisher
Corrections
All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:igg:jswis0:v:22:y:2026:i:1:p:1-35. See general information about how to correct material in RePEc.
If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.
We have no bibliographic references for this item. You can help adding them by using this form .
If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.
For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: Journal Editor (email available below). General contact details of provider: https://www.igi-global.com .
Please note that corrections may take a couple of weeks to filter through
the various RePEc services.