用户名: 密码: 验证码:
Computational Modeling and Psychophysics in Low- and Mid-Level Vision.
详细信息   
  • 作者:Hou ; Xiaodi.
  • 学历:Ph.D.
  • 年:2014
  • 毕业院校:California Institute of Technology
  • Department:Computation and Neural Systems
  • ISBN:9781303946592
  • CBH:3622737
  • Country:USA
  • 语种:English
  • FileSize:50142147
  • Pages:126
文摘
This thesis addresses a series of topics related to the question of how people find the foreground objects from complex scenes. With both computer vision modeling,as well as psychophysical analyses,we explore the computational principles for low- and mid-level vision. We first explore the computational methods of generating saliency maps from images and image sequences. We propose an extremely fast algorithm called Image Signature that detects the locations in the image that attract human eye gazes. With a series of experimental validations based on human behavioral data collected from various psychophysical experiments,we conclude that the Image Signature and its spatial-temporal extension,the Phase Discrepancy,are among the most accurate algorithms for saliency detection under various conditions. In the second part,we bridge the gap between fixation prediction and salient object segmentation with two efforts. First,we propose a new dataset that contains both fixation and object segmentation information. By simultaneously presenting the two types of human data in the same dataset,we are able to analyze their intrinsic connection,as well as understanding the drawbacks of todays "standard" but inappropriately labeled salient object segmentation dataset. Second,we also propose an algorithm of salient object segmentation. Based on our novel discoveries on the connections of fixation data and salient object segmentation data,our model significantly outperforms all existing models on all 3 datasets with large margins. In the third part of the thesis,we discuss topics around the human factors of boundary analysis. Closely related to salient object segmentation,boundary analysis focuses on delimiting the local contours of an object. We identify the potential pitfalls of algorithm evaluation for the problem of boundary detection. Our analysis indicates that todays popular boundary detection datasets contain significant level of noise,which may severely influence the benchmarking results. To give further insights on the labeling process,we propose a model to characterize the principles of the human factors during the labeling process. The analyses reported in this thesis offer new perspectives to a series of interrelating issues in low- and mid-level vision. It gives warning signs to some of todays "standard" procedures,while proposing new directions to encourage future research.

© 2004-2018 中国地质图书馆版权所有 京ICP备05064691号 京公网安备11010802017129号

地址:北京市海淀区学院路29号 邮编:100083

电话:办公室:(+86 10)66554848;文献借阅、咨询服务、科技查新:66554700