How to detect meaningful video representation becomes an interesting topic in various research communities. Visual Attention System is proposed to detect “Region of Interesting” from input video sequence. Generally the attended regions correspond ...
How to detect meaningful video representation becomes an interesting topic in various research communities. Visual Attention System is proposed to detect “Region of Interesting” from input video sequence. Generally the attended regions correspond to visually prominent object in the image in video sequence.
In this paper, we suggest a novel method to improve previous approaches based on spatiotemporal attention modules. We propose a new integration model to make use of depth information in addition to spatiotemporal features. As a result the proposed method compensates typical approaches for inaccuracy improving performance.
Motion is important cue when we derive temporal saliency. But noise obtained during the input and computation process deteriorates accuracy of temporal saliency. To remove the noise from motion information we exploited the result of psychological studies. The function of "double opponent receptive field" and "noise filtration" in Middle Temporal area is simulated. We also applied "FlagMap" on each frame to prevent "Flickering" of global-area noise. These considerations make the proposed system detect the salient regions with higher accuracy while removing noise effectively.
Depth information is used to detect salient regions. Typical systems get problems in determining the saliency if several salient regions are partially occluded and/or have almost equal saliency. However, the proposed system is able to separate the regions with high accuracy. Spatiotemporally separated prominent regions in the first stage are prioritized using depth value one by one in the second stage.
The proposed method has been applied to several image sequences. Experiment result shows that the proposed method can describe the salient regions with higher accuracy than the previous approaches do.