Computational Modeling and Psychophysics in Low- and Mid-Level Vision.

设为首页

收藏本站

网站地图 | English | 公务邮箱

远程访问

NSTL服务站

Computational Modeling and Psychophysics in Low- and Mid-Level Vision.

详细信息

作者：Hou ; Xiaodi.
学历：Ph.D.
年：2014
毕业院校：California Institute of Technology
Department：Computation and Neural Systems
ISBN：9781303946592
CBH：3622737
Country：USA
语种：English
FileSize：50142147
Pages：126

文摘

This thesis addresses a series of topics related to the question of how people find the foreground objects from complex scenes. With both computer vision modeling,as well as psychophysical analyses,we explore the computational principles for low- and mid-level vision. We first explore the computational methods of generating saliency maps from images and image sequences. We propose an extremely fast algorithm called Image Signature that detects the locations in the image that attract human eye gazes. With a series of experimental validations based on human behavioral data collected from various psychophysical experiments,we conclude that the Image Signature and its spatial-temporal extension,the Phase Discrepancy,are among the most accurate algorithms for saliency detection under various conditions. In the second part,we bridge the gap between fixation prediction and salient object segmentation with two efforts. First,we propose a new dataset that contains both fixation and object segmentation information. By simultaneously presenting the two types of human data in the same dataset,we are able to analyze their intrinsic connection,as well as understanding the drawbacks of todays "standard" but inappropriately labeled salient object segmentation dataset. Second,we also propose an algorithm of salient object segmentation. Based on our novel discoveries on the connections of fixation data and salient object segmentation data,our model significantly outperforms all existing models on all 3 datasets with large margins. In the third part of the thesis,we discuss topics around the human factors of boundary analysis. Closely related to salient object segmentation,boundary analysis focuses on delimiting the local contours of an object. We identify the potential pitfalls of algorithm evaluation for the problem of boundary detection. Our analysis indicates that todays popular boundary detection datasets contain significant level of noise,which may severely influence the benchmarking results. To give further insights on the labeling process,we propose a model to characterize the principles of the human factors during the labeling process. The analyses reported in this thesis offer new perspectives to a series of interrelating issues in low- and mid-level vision. It gives warning signs to some of todays "standard" procedures,while proposing new directions to encourage future research.

地址：北京市海淀区学院路29号邮编：100083

电话：办公室：(+86 10)66554848；文献借阅、咨询服务、科技查新：66554700