Abstract:Visual attention is one of the most significant characteristics for selecting and understanding the outside redundancy world. The human vision system cannot process all information simultaneously due to the visual information bottleneck. In order to reduce the redundant input of visual information, the human visual system mainly focuses on dominant parts of scenes. This is commonly known as visual saliency map prediction. This paper proposed a new psychophysical saliency prediction architecture, WECSF, inspired by multi-channel model of visual cortex functioning in humans. The model consists of opponent color channels, wavelet transform, wavelet energy map, and contrast sensitivity function for extracting low-level image features and providing a maximum approximation to the human visual system. The proposed model is evaluated using several datasets, including the MIT1003, MIT300, TORONTO, SID4VAM, and UCF Sports datasets. We also quantitatively and qualitatively compare the saliency prediction performance with that of other state-of-the-art models. Our model achieved strongly stable and better performance with different metrics on natural images, psychophysical synthetic images and dynamic videos. Additionally, we found that Fourier and spectral-inspired saliency prediction models outperformed other state-of-the-art non-neural network and even deep neural network models on psychophysical synthetic images. It can be explained and supported by the Fourier Vision Hypothesis. In the meantime, we suggest that deep neural networks need specific architectures and goals to be able to predict salient performance on psychophysical synthetic images better and more reliably. Finally, the proposed model could be used as a computational model of primate vision system and help us understand mechanism of primate vision system.

A Computational Model for Stereoscopic Visual Saliency Prediction

Learning Stereoscopic Visual Attention Model for 3d Video

Stereoscopic visual saliency prediction based on stereo contrast and stereo focus

A Novel Saliency Model for Stereoscopic Images.

Saliency Detection for Stereoscopic Images Based on Depth Confidence Analysis and Multiple Cues Fusion

A Learning-Based Visual Saliency Prediction Model for Stereoscopic 3D Video (LBVS-3D)

A Three-Pathway Psychobiological Framework of Salient Object Detection Using Stereoscopic Technology.

Salient Object Detection and Classification for Stereoscopic Images

Stereoscopic Image Quality Assessment Method Based on Binocular Combination Saliency Model

A Psychophysically Oriented Saliency Map Prediction Model

A biologically inspired computational model for image saliency detection.

Pre-attentive segmentation and correspondence in stereo.

Pre–attentive Segmentation and Correspondence in Stereo

Stereoscopic video conversion based on depth tracking

Toward a Quality Predictor for Stereoscopic Images via Analysis of Human Binocular Visual Perception

Depth saliency based on anisotropic center-surround difference

Predicting human gaze beyond pixels.

Automatic Tag Saliency Ranking for Stereo Images

Video Saliency Detection via Dynamic Consistent Spatio-Temporal Attention Modelling.

Quality Assessment Metric of Stereo Images Considering Cyclopean Integration and Visual Saliency.

SAL3D: a model for saliency prediction in 3D meshes