DG-SegNet: A depth-guided RGB-D semantic segmentation framework for emergency escape ramps toward traffic accident prevention

OBJECTIVE: In intelligent transportation scenarios, semantic segmentation plays a crucial role in autonomous driving perception by enabling fine-grained semantic understanding of complex road environments and contributing to the effective reduction of traffic accident risks. However, conventional RGB-based segmentation methods exhibit inherent limitations, and although multimodal information fusion has emerged as a promising direction, existing multimodal models generally suffer from large parameter sizes and high computational complexity. METHODS: To address the aforementioned challenges, we propose DG-SegNet, an efficient depth-guided RGB-D semantic segmentation framework based on SegFormer. The proposed method introduces depth information exclusively at the shallow C1 feature stage, thereby reducing redundant cross-modal computation, and incorporates a Multi-scale Feature Refinement module together with a Gated Fusion mechanism to structurally reorganize heterogeneous modal features and adaptively regulate cross-modal information flow. The fused multimodal features are subsequently unified to enhance the coherence of the overall feature representation. The dataset was expanded from 411 to 500 samples, with the number of semantic categories increased to nine. RESULT: Our method achieves an MIoU of 78.77% and an MPA of 84.89%, outperforming six classical and state-of-the-art competing approaches evaluated under the same experimental settings. Furthermore, on the public PST900 dataset, comparative experiments against eight advanced methods demonstrate that DG-SegNet attains an MIoU of 84.21% and an MPA of 88.56%, consistently maintaining superior performance across multiple semantic categories. CONCLUSIONS: DG-SegNet mitigates the semantic representation limitations of single RGB inputs in complex traffic scenarios and provides a feasible solution that balances accuracy and efficiency for effective environmental perception in autonomous driving and intelligent transportation systems, demonstrating strong practical applicability and promising potential for real-world deployment.

Authors

Institutions

Publication Details

Journal
Traffic Injury Prevention
Published
2026-09-14
DOI
https://doi.org/10.1080/15389588.2026.2713113
Primary Topic
Advanced Neural Network Applications
Type
article
Field-Weighted Citation Impact
0.00
Controls
|||
ALL TIME
JAN
FEB
MAR
APR
MAY
JUN
JUL
AUG
SEP
article

DG-SegNet: A depth-guided RGB-D semantic segmentation framework for emergency escape ramps toward traffic accident prevention

Zuosheng Hu, Yun Hao, Luping Wang, Guiling Li et al.
Traffic Injury Prevention
Advanced Neural Network Applications
article

DG-SegNet: A depth-guided RGB-D semantic segmentation framework for emergency escape ramps toward traffic accident prevention

Zuosheng Hu, Yun Hao, Luping Wang, Guiling Li, Junhao Li
article en

Abstract

OBJECTIVE: In intelligent transportation scenarios, semantic segmentation plays a crucial role in autonomous driving perception by enabling fine-grained semantic understanding of complex road environments and contributing to the effective reduction of traffic accident risks. However, conventional RGB-based segmentation methods exhibit inherent limitations, and although multimodal information fusion has emerged as a promising direction, existing multimodal models generally suffer from large parameter sizes and high computational complexity. METHODS: To address the aforementioned challenges, we propose DG-SegNet, an efficient depth-guided RGB-D semantic segmentation framework based on SegFormer. The proposed method introduces depth information exclusively at the shallow C1 feature stage, thereby reducing redundant cross-modal computation, and incorporates a Multi-scale Feature Refinement module together with a Gated Fusion mechanism to structurally reorganize heterogeneous modal features and adaptively regulate cross-modal information flow. The fused multimodal features are subsequently unified to enhance the coherence of the overall feature representation. The dataset was expanded from 411 to 500 samples, with the number of semantic categories increased to nine. RESULT: Our method achieves an MIoU of 78.77% and an MPA of 84.89%, outperforming six classical and state-of-the-art competing approaches evaluated under the same experimental settings. Furthermore, on the public PST900 dataset, comparative experiments against eight advanced methods demonstrate that DG-SegNet attains an MIoU of 84.21% and an MPA of 88.56%, consistently maintaining superior performance across multiple semantic categories. CONCLUSIONS: DG-SegNet mitigates the semantic representation limitations of single RGB inputs in complex traffic scenarios and provides a feasible solution that balances accuracy and efficiency for effective environmental perception in autonomous driving and intelligent transportation systems, demonstrating strong practical applicability and promising potential for real-world deployment.

Traffic Injury Prevention
University of Shanghai for Science and Technology (CN), Southwest Jiaotong University (CN), Henan Normal University (CN)
Openalex Percentile: Top 13%
Advanced Neural Network Applications
AI Navigator

Ask Laika to Summarize, Analyze, and Connect papers live on the map.

Summarize Papers & Methodologies

Extract key findings, datasets, and comparative methods across publications.

Benchmark Rankings & Visual Analytics

Rank top research institutions, authors, funders, topics, and journals by Field-Weighted Citation Impact (FWCI) and paper volume with instant charts.

Connect Distant Disciplines

Bridge topological clusters on the map to find hidden collaborative intersections.