English

Q-YOLOP: Quantization-aware You Only Look Once for Panoptic Driving Perception

Computer Vision and Pattern Recognition 2023-07-11 v1 Artificial Intelligence

Abstract

In this work, we present an efficient and quantization-aware panoptic driving perception model (Q- YOLOP) for object detection, drivable area segmentation, and lane line segmentation, in the context of autonomous driving. Our model employs the Efficient Layer Aggregation Network (ELAN) as its backbone and task-specific heads for each task. We employ a four-stage training process that includes pretraining on the BDD100K dataset, finetuning on both the BDD100K and iVS datasets, and quantization-aware training (QAT) on BDD100K. During the training process, we use powerful data augmentation techniques, such as random perspective and mosaic, and train the model on a combination of the BDD100K and iVS datasets. Both strategies enhance the model's generalization capabilities. The proposed model achieves state-of-the-art performance with an mAP@0.5 of 0.622 for object detection and an mIoU of 0.612 for segmentation, while maintaining low computational and memory requirements.

Keywords

Cite

@article{arxiv.2307.04537,
  title  = {Q-YOLOP: Quantization-aware You Only Look Once for Panoptic Driving Perception},
  author = {Chi-Chih Chang and Wei-Cheng Lin and Pei-Shuo Wang and Sheng-Feng Yu and Yu-Chen Lu and Kuan-Cheng Lin and Kai-Chiang Wu},
  journal= {arXiv preprint arXiv:2307.04537},
  year   = {2023}
}
R2 v1 2026-06-28T11:25:56.329Z