L-CoIns: Language-Based Colorization with Instance Awareness

Zheng Chang,Shuchen Weng,Peixuan Zhang,Yu Li,Si Li,Boxin Shi
DOI: https://doi.org/10.1109/cvpr52729.2023.01842
2023-01-01
Abstract:Language-based colorization produces plausible colors consistent with the language description provided by the user. Recent studies introduce additional annotation to prevent color-object coupling and mismatch issues, but they still have difficulty in distinguishing instances corresponding to the same object words. In this paper, we propose a transformer-based framework to automatically aggregate similar image patches and achieve instance awareness without any additional knowledge. By applying our presented luminance augmentation and counter-color loss to break down the statistical correlation between luminance and color words, our model is driven to synthesize colors with better descriptive consistency. We further collect a dataset to provide distinctive visual characteristics and detailed language descriptions for multiple instances in the same image. Extensive experiments demonstrate our advantages of synthesizing visually pleasing and description-consistent results of instance-aware colorization.
What problem does this paper attempt to address?