3D-GMNet: Learning to Estimate 3D Shape from A Single Image As A Gaussian Mixture

Yamashita, Kohei; Nobuhara, Shohei; Nishino, Ko

Computer Science > Computer Vision and Pattern Recognition

arXiv:1912.04663v1 (cs)

[Submitted on 10 Dec 2019 (this version), latest version 15 Aug 2020 (v2)]

Title:3D-GMNet: Learning to Estimate 3D Shape from A Single Image As A Gaussian Mixture

Authors:Kohei Yamashita, Shohei Nobuhara, Ko Nishino

View PDF

Abstract:In this paper, we introduce 3D-GMNet, a deep neural network for single-image 3D shape recovery. As the name suggests, 3D-GMNet recovers 3D shape as a Gaussian mixture model. In contrast to voxels, point clouds, or meshes, a Gaussian mixture representation requires a much smaller footprint for representing 3D shapes and, at the same time, offers a number of additional advantages including instant pose estimation, automatic level-of-detail computation, and a distance measure. The proposed 3D-GMNet is trained end-to-end with single input images and corresponding 3D models by using two novel loss functions: a 3D Gaussian mixture loss and a multi-view 2D loss. The first maximizes the likelihood of the Gaussian mixture shape representation by considering the target point cloud as samples from the true distribution, and the latter improves the consistency between the input silhouette and the projection of the Gaussian mixture shape model. Extensive quantitative evaluations with synthesized and real images demonstrate the effectiveness of the proposed method.

Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:1912.04663 [cs.CV]
	(or arXiv:1912.04663v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.1912.04663

Submission history

From: Shohei Nobuhara [view email]
[v1] Tue, 10 Dec 2019 12:23:24 UTC (5,765 KB)
[v2] Sat, 15 Aug 2020 14:18:31 UTC (21,871 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:3D-GMNet: Learning to Estimate 3D Shape from A Single Image As A Gaussian Mixture

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:3D-GMNet: Learning to Estimate 3D Shape from A Single Image As A Gaussian Mixture

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators