Learning a generative model of images by factoring appearance and shape
Neural Computation(2011)
摘要
Computer vision has grown tremendously in the past two decades. Despite all efforts, existing attempts at matching parts of the human visual system's extraordinary ability to understand visual scenes lack either scope or power. By combining the advantages of general low-level generative models and powerful layer-based and hierarchical models, this work aims at being a first step toward richer, more flexible models of images. After comparing various types of restricted Boltzmann machines (RBMs) able to model continuous-valued data, we introduce our basic model, the masked RBM, which explicitly models occlusion boundaries in image patches by factoring the appearance of any patch region from its shape. We then propose a generative model of larger images using a field of such RBMs. Finally, we discuss how masked RBMs could be stacked to form a deep model able to generate more complicated structures and suitable for various tasks such as segmentation or object recognition.
更多查看译文
关键词
basic model,deep model,flexible model,general low-level generative model,generative model,hierarchical model,models occlusion boundary,human visual system,various task,various type
AI 理解论文
溯源树
样例
生成溯源树,研究论文发展脉络