AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
PDF (12 MB)
Collect
Submit Manuscript AI Chat Paper
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Research Article | Open Access

DeepPrimitive: Image decomposition by layered primitive detection

Tsinghua University, Beijing, 100084, China.
Computer Science Department, University of Toronto, Toronto, M5S2E4, Canada.
Stanford University, Stanford, 94305, United States.
University of California San Diego, La Jolla, 92093, United States.
University of Wisconsin-Madison, Madison, 53715, United States.
Show Author Information

Abstract

The perception of the visual world through basic building blocks, such as cubes, spheres, and cones, gives human beings a parsimonious understanding of the visual world. Thus, efforts to find primitive-based geometric interpretations of visual data date back to 1970s studies of visual media. However, due to the difficulty of primitive fitting in the pre-deep learning age, this research approach faded from the main stage, and the vision community turned primarily to semantic image understanding. In this paper, we revisit the classical problem of building geometric interpretations of images, using supervised deep learning tools. We build a framework to detect primitives from images in a layered manner by modifying the YOLO network; an RNN with a novel loss function is then used to equip this network with the capability to predict primitives with a variable number of parameters. We compare our pipeline to traditional and other baseline learning methods, demonstrating that our layered detection model has higher accuracy and performs better reconstruction.

Electronic Supplementary Material

Download File(s)
41095_2018_128_MOESM1_ESM.pdf (8.1 MB)

References

【1】
【1】
 
 
Computational Visual Media
Pages 385-397

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
Huang J, Gao J, Ganapathi-Subramanian V, et al. DeepPrimitive: Image decomposition by layered primitive detection. Computational Visual Media, 2018, 4(4): 385-397. https://doi.org/10.1007/s41095-018-0128-6

1663

Views

90

Downloads

5

Crossref

N/A

Web of Science

6

Scopus

0

CSCD

Revised: 30 November 2018
Accepted: 03 December 2018
Published: 23 December 2018
© The author(s) 2018

This article is published with open access at Springerlink.com

The articles published in this journal are distributed under the terms of the Creative Commons Attribution 4.0 International License (http://creativecommons.org/licenses/by/4.0/), which permits unrestricted use, distribution, and reproduction in any medium, provided you give appropriate credit to the original author(s) and the source, provide a link to the Creative Commons license, and indicate if changes were made.

Other papers from this open access journal are available free of charge from http://www.springer.com/journal/41095. To submit a manuscript, please go to https://www.editorialmanager.com/cvmj.