<script type="text/javascript">
<!--
document.write('<div id="oa_widget"></div>');
document.write('<script type="text/javascript" src="https://www.openaire.eu/index.php?option=com_openaire&view=widget&format=raw&projectId=undefined&type=result"></script>');
-->
</script>

COPY SCRIPT

For further information contact us at helpdesk@openaire.eu

Style-Hallucinated Dual Consistency Learning: A Unified Framework for Visual Domain Generalization

descriptionPublicationkeyboard_double_arrow_right Article , Preprint , Other literature type 18 Oct 2023Embargo end date: 01 Jan 2022 Italy Publisher:Springer Science and Business Media LLCJournal:International Journal of Computer Vision, volume 132, pages 837-853 (issn: 0920-5691, eissn: 1573-1405,

Authors: Zhao, Yuyang; Zhong, Zhun; Zhao, Na; Sebe, Nicu; Lee, Gim Hee;

doi: 10.1007/s11263-023-01911-w , 10.48550/arxiv.2212.09068

arXiv: http://arxiv.org/abs/2212.09068

handle: 11572/404536

Style-Hallucinated Dual Consistency Learning: A Unified Framework for Visual Domain Generalization

- Summary
- Subjects
- Related research
  (2)
- Metrics

Abstract

Domain shift widely exists in the visual world, while modern deep neural networks commonly suffer from severe performance degradation under domain shift due to the poor generalization ability, which limits the real-world applications. The domain shift mainly lies in the limited source environmental variations and the large distribution gap between source and unseen target data. To this end, we propose a unified framework, Style-HAllucinated Dual consistEncy learning (SHADE), to handle such domain shift in various visual tasks. Specifically, SHADE is constructed based on two consistency constraints, Style Consistency (SC) and Retrospection Consistency (RC). SC enriches the source situations and encourages the model to learn consistent representation across style-diversified samples. RC leverages general visual knowledge to prevent the model from overfitting to source data and thus largely keeps the representation consistent between the source and general visual models. Furthermore, we present a novel style hallucination module (SHM) to generate style-diversified samples that are essential to consistency learning. SHM selects basis styles from the source distribution, enabling the model to dynamically generate diverse and realistic samples during training. Extensive experiments demonstrate that our versatile SHADE can significantly enhance the generalization in various visual recognition tasks, including image classification, semantic segmentation and object detection, with different models, i.e., ConvNets and Transformer.

Accepted by IJCV. Journal extension of arXiv:2204.02548. Code is available at https://github.com/HeliosZhao/SHADE-VisualDG

Country

Italy

Related Organizations

Nottingham Trent University
United Kingdom
National University of Singapore
Singapore
University of Trento
Italy
Singapore University of Technology and Design
Singapore

Keywords

FOS: Computer and information sciences, Consistency learning; Domain generalization; Style variation; Visual recognition, Computer Vision and Pattern Recognition (cs.CV), Computer Science - Computer Vision and Pattern Recognition

2 Research products, page 1 of 1

Single-DGOD software on GitHub
IsRelatedTo
SHADE-VisualDG software on GitHub
IsRelatedTo

Impact byBIP!

	citations This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	17
	popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network.	Top 10%
	influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	Average
	impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network.	Top 10%