arrow
返回

Learning Direct Optimization for scene understanding

delete2020-09-01
delete2
delete
OA
AI
C
Christopher K. I. Williams
J
John Winn
DOI:10.1016/j.patcog.2020.107369delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
We develop a Learning Direct Optimization (LiDO) method for the refinement of a latent variable model that describes input image x. Our goal is to explain a single image x with an interpretable 3D computer graphics model having scene graph latent variables z (such as object appearance, camera position). Given a current estimate of z we can render a prediction of the image g(z), which can be compared to the image x. The standard way to proceed is then to measure the error E(x, g(z)) between the two, and use an optimizer to minimize the error. However, it is unknown which error measure E would be most effective for simultaneously addressing issues such as misaligned objects, occlusions, textures, etc. In contrast, the LiDO approach trains a Prediction Network to predict an update directly to correct z, rather than minimizing the error with respect to z. Experiments show that LiDO converges rapidly as it does not need to perform a search on the error landscape, produces better solutions than error-based competitors, and is able to handle the mismatch between the data and the fitted scene model. We apply LiDO to a realistic synthetic dataset, and show that the method also transfers to work well with real images. (C) 2020 Elsevier Ltd. All rights reserved.
Keyword:
Computer vision
Scene understanding
3D Reconstruction
Inverse graphics
Object recognition
Scene graph
Analysis-by-synthesis
Graphics
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Pattern Recognition 封面图
Pattern Recognition
IF:
7.6
论文数:
1.3W
被引数:
4.5W

机构

U
University of Edinburgh
学者数:
5.2W
论文数: 4.6W
被引数: 71
引用论文

引用论文

err分享
err收藏
Selective Search for Object Recognition
err2013-04-02
err3.9K
PREAI
errUijlings, J. R. R.; van de Sande, K. E. A.; Gevers, T.; Smeulders, A. W. M.
err分享
err收藏
The three R's of computer vision: Recognition, reconstruction and reorganization
err2016-03-01
err20
errOAAI
errMalik, Jitendra; Arbelaez, Pablo; Carreira, Joao; Fragkiadaki, Katerina; Girshick, Ross; Gkioxari, Georgia; Gupta, Saurabh; Hariharan, Bharath; Kar, Abhishek; Tulsiani, Shubham
err分享
err收藏
Electrochemically assisted micro localized grafting of aptamers in a microchannel engraved in fluorinated thermoplastic polymer Dyneon THV
err2015-01-01
err0
PREAI
errC. Perréard; Y. Ladner; F. d'Orlyé; S. Descroix; V. Taniga; A. Varenne; F. Kanoufi; C. Slim; S. Griveau; F. Bedioui
err分享
err收藏
CRF learning with CNN features for image segmentation
err2015-10-01
err177
errOAAI
errLiu, Fayao; Lin, Guosheng; Shen, Chunhua
err分享
err收藏
ImageNet Large Scale Visual Recognition ChallengeImageNet大规模视觉识别挑战
err2015-04-11
err2.7W
PREAI
errRussakovsky, Olga; Deng, Jia; Su, Hao; Krause, Jonathan; Satheesh, Sanjeev; Ma, Sean; Huang, Zhiheng; Karpathy, Andrej; Khosla, Aditya; Bernstein, Michael; Berg, Alexander C.; Fei-Fei, Li
err分享
err收藏
Indoor Scene Understanding with Geometric and Semantic Contexts
err2014-11-12
err34
PREAI
errChoi, Wongun; Chao, Yu-Wei; Pantofaru, Caroline; Savarese, Silvio
err分享
err收藏
Putting objects in perspective
err2008-04-17
err352
PREAI
errHoiem, Derek; Efros, Alexei A.; Hebert, Martial
err分享
err收藏
Complete 3D Scene Parsing from an RGBD Image
err2018-11-21
err18
PREAI
errZou, Chuhang; Guo, Ruiqi; Li, Zhizhong; Hoiem, Derek
err分享
err收藏
学者 查看更多内容