ST3D++: Denoised Self-Training for Unsupervised Domain Adaptation on 3D Object Detection

Yang, Jihan; Shi, Shaoshuai; Wang, Zhe; Li, Hongsheng; Qi, Xiaojuan

File Download

There are no files associated with this item.

Links for fulltext

(May Require Subscription)

Publisher Website: 10.1109/TPAMI.2022.3216606
WOS: WOS:000964792800065
Find via

Supplementary

Citations:
- Web of Science: 0
Appears in Collections:
- Electrical & Electronic Engineering: Journal/Magazine Articles

Article: ST3D++: Denoised Self-Training for Unsupervised Domain Adaptation on 3D Object Detection

Title	ST3D++: Denoised Self-Training for Unsupervised Domain Adaptation on 3D Object Detection
Authors	Yang, Jihan Shi, Shaoshuai Wang, Zhe Li, Hongsheng Qi, Xiaojuan
Issue Date	25-Oct-2022
Publisher	Institute of Electrical and Electronics Engineers
Citation	IEEE Transactions on Pattern Analysis and Machine Intelligence, 2022, v. 45, n. 5, p. 6354-6371 How to Cite? DOI: http://dx.doi.org/10.1109/TPAMI.2022.3216606
Abstract	In this paper, we present a self-training method, named ST3D++, with a holistic pseudo label denoising pipeline for unsupervised domain adaptation on 3D object detection. ST3D++ aims at reducing noise in pseudo label generation as well as alleviating the negative impacts of noisy pseudo labels on model training. First, ST3D++ pre-trains the 3D object detector on the labeled source domain with random object scaling (ROS) which is designed to reduce target domain pseudo label noise arising from object scale bias of the source domain. Then, the detector is progressively improved through alternating between generating pseudo labels and training the object detector with pseudo-labeled target domain data. Here, we equip the pseudo label generation process with a hybrid quality-aware triplet memory to improve the quality and stability of generated pseudo labels. Meanwhile, in the model training stage, we propose a source data assisted training strategy and a curriculum data augmentation policy to effectively rectify noisy gradient directions and avoid model over-fitting to noisy pseudo labeled data. These specific designs enable the detector to be trained on meticulously refined pseudo labeled target data with denoised training signals, and thus effectively facilitate adapting an object detector to a target domain without requiring annotations. Finally, our method is assessed on four 3D benchmark datasets (i.e., Waymo, KITTI, Lyft, and nuScenes) for three common categories (i.e., car, pedestrian and bicycle). ST3D++ achieves state-of-the-art performance on all evaluated settings, outperforming the corresponding baseline by a large margin (e.g., 9.6% ∼ 38.16% on Waymo → KITTI in terms of AP 3D ), and even surpasses the fully supervised oracle results on the KITTI 3D object detection benchmark with target prior. Code is available at https://github.com/CVMI-Lab/ST3D .
Persistent Identifier	http://hdl.handle.net/10722/328497
ISSN	0162-8828 2023 Impact Factor: 20.8 2023 SCImago Journal Rankings: 6.158
ISI Accession Number ID	WOS:000964792800065

DC Field	Value	Language
dc.contributor.author	Yang, Jihan	-
dc.contributor.author	Shi, Shaoshuai	-
dc.contributor.author	Wang, Zhe	-
dc.contributor.author	Li, Hongsheng	-
dc.contributor.author	Qi, Xiaojuan	-
dc.date.accessioned	2023-06-28T04:45:30Z	-
dc.date.available	2023-06-28T04:45:30Z	-
dc.date.issued	2022-10-25	-
dc.identifier.citation	IEEE Transactions on Pattern Analysis and Machine Intelligence, 2022, v. 45, n. 5, p. 6354-6371	-
dc.identifier.issn	0162-8828	-
dc.identifier.uri	http://hdl.handle.net/10722/328497	-
dc.description.abstract	<p>In this paper, we present a self-training method, named ST3D++, with a holistic pseudo label denoising pipeline for unsupervised domain adaptation on 3D object detection. ST3D++ aims at reducing noise in pseudo label generation as well as alleviating the negative impacts of noisy pseudo labels on model training. First, ST3D++ pre-trains the 3D object detector on the labeled source domain with random object scaling (ROS) which is designed to reduce target domain pseudo label noise arising from object scale bias of the source domain. Then, the detector is progressively improved through alternating between generating pseudo labels and training the object detector with pseudo-labeled target domain data. Here, we equip the pseudo label generation process with a hybrid quality-aware triplet memory to improve the quality and stability of generated pseudo labels. Meanwhile, in the model training stage, we propose a source data assisted training strategy and a curriculum data augmentation policy to effectively rectify noisy gradient directions and avoid model over-fitting to noisy pseudo labeled data. These specific designs enable the detector to be trained on meticulously refined pseudo labeled target data with denoised training signals, and thus effectively facilitate adapting an object detector to a target domain without requiring annotations. Finally, our method is assessed on four 3D benchmark datasets (i.e., Waymo, KITTI, Lyft, and nuScenes) for three common categories (i.e., car, pedestrian and bicycle). ST3D++ achieves state-of-the-art performance on all evaluated settings, outperforming the corresponding baseline by a large margin (e.g., 9.6% ∼ 38.16% on Waymo → KITTI in terms of AP 3D ), and even surpasses the fully supervised oracle results on the KITTI 3D object detection benchmark with target prior. Code is available at https://github.com/CVMI-Lab/ST3D .<br></p>	-
dc.language	eng	-
dc.publisher	Institute of Electrical and Electronics Engineers	-
dc.relation.ispartof	IEEE Transactions on Pattern Analysis and Machine Intelligence	-
dc.title	ST3D++: Denoised Self-Training for Unsupervised Domain Adaptation on 3D Object Detection	-
dc.type	Article	-
dc.identifier.doi	10.1109/TPAMI.2022.3216606	-
dc.identifier.volume	45	-
dc.identifier.issue	5	-
dc.identifier.spage	6354	-
dc.identifier.epage	6371	-
dc.identifier.eissn	1939-3539	-
dc.identifier.isi	WOS:000964792800065	-
dc.identifier.issnl	0162-8828	-

File Download

Links for fulltext

(May Require Subscription)

Supplementary

Article: ST3D++: Denoised Self-Training for Unsupervised Domain Adaptation on 3D Object Detection

Export via OAI-PMH Interface in XML Formats

OR

Export to Other Non-XML Formats