Adversarial Semantic Scene Completion from a Single Depth Image

Authors

Yida Wang, David Tan, Nassir Navab and Federico Tombari

International Conference on 3D Vision, IEEE

Showcase

Overview

We introduce a direct reconstruction method to reconstruct from a 2.5D depth image to a 3D voxel data with both shape completion and semantic segmentation that relies on a deep architecture based on 3D VAE with an adversarial training to improve the performance of this task.

Architecture

We utilize the latent representation of 3D auto-encoder to help train a latent representation from a depth image. The 3D auto-encoder is removed after the parametric model is trained. This pipeline is optimized for the encoders for the depth image and the 3D volumetric data and the shared generator is also optimised during training.

Discriminators

To make the latent representation and the reconstructed 3D scene similar to each others, we apply two discriminators for both targets. In this manner, the latent representation of the depth produces the expected target more precisely compared to the latent representation of the ground truth volumetric data.

Name		Name	Last commit message	Last commit date
Latest commit History 16 Commits
3dv		3dv
data		data
depth-tsdf		depth-tsdf
visualization		visualization
.DS_Store		.DS_Store
README.md		README.md
config.py		config.py
config_test.py		config_test.py
evaluate.py		evaluate.py
main.py		main.py
model.py		model.py
train.py		train.py
tsdf.ply		tsdf.ply
util.py		util.py

dtmfgold/gan-depth-semantic3d

Folders and files

Latest commit

History

Repository files navigation

Adversarial Semantic Scene Completion from a Single Depth Image

Authors

Showcase

Overview

Architecture

Discriminators

Our data format

Qualitative results

About

Resources

Stars

Watchers

Forks

Languages