MINERVAS: Massive INterior EnviRonments VirtuAl Synthesis-Reference-Cited by-同舟云学术

MINERVAS: Massive INterior EnviRonments VirtuAl Synthesis

Published:2022-10 Issue:7 Volume:41 Page:63-74
ISSN:0167-7055
Container-title:Computer Graphics Forum
language:en
Short-container-title:Computer Graphics Forum

Author:

Ren Haocheng¹,Zhang Hao¹,Zheng Jia²,Zheng Jiaxiang²,Tang Rui²,Huo Yuchi¹,Bao Hujun¹,Wang Rui¹

Affiliation:

1. State Key Lab of CAD&CG Zhejiang University

2. Manycore Tech Inc.

Abstract

AbstractWith the rapid development of data‐driven techniques, data has played an essential role in various computer vision tasks. Many realistic and synthetic datasets have been proposed to address different problems. However, there are lots of unresolved challenges: (1) the creation of dataset is usually a tedious process with manual annotations, (2) most datasets are only designed for a single specific task, (3) the modification or randomization of the 3D scene is difficult, and (4) the release of commercial 3D data may encounter copyright issue. This paper presents MINERVAS, a Massive INterior EnviRonments VirtuAl Synthesis system, to facilitate the 3D scene modification and the 2D image synthesis for various vision tasks. In particular, we design a programmable pipeline with Domain‐Specific Language, allowing users to select scenes from the commercial indoor scene database, synthesize scenes for different tasks with customized rules, and render various types of imagery data, such as color images, geometric structures, semantic labels. Our system eases the difficulty of customizing massive scenes for different tasks and relieves users from manipulating fine‐grained scene configurations by providing user‐controllable randomness using multilevel samplers. Most importantly, it empowers users to access commercial scene databases with millions of indoor scenes and protects the copyright of core data assets, e.g., 3D CAD models. We demonstrate the validity and flexibility of our system by using our synthesized data to improve the performance on different kinds of computer vision tasks. The project page is at https://coohom.github.io/MINERVAS.

Funder

National Natural Science Foundation of China-Liaoning Joint Fund

Fundamental Research Funds for the Central Universities

Publisher

Wiley

Subject

Computer Graphics and Computer-Aided Design

Link

https://onlinelibrary.wiley.com/doi/pdf/10.1111/cgf.14657

Reference78 articles.

1. AvetisyanA. DahnertM. DaiA. SavvaM. ChangA. X. NiessnerM.: Scan2cad: Learning cad model alignment in rgb-d scans. InCVPR(2019) pp.2614–2623. 2

2. ArmeniI. SenerO. ZamirA. R. JiangH. BrilakisI. FischerM. SavareseS.: 3d semantic parsing of large-scale indoor spaces. InCVPR(2016) pp.1534–1543. 2

3. ArmeniI. SaxS. ZamirA. R. SavareseS.: Joint 2d-3d-semantic data for indoor scene understanding.CoRR abs/1702.01105(2017). 2 9

4. BhatS. F. AlhashimI. WonkaP.: Adabins: Depth estimation using adaptive bins. InCVPR(2021) pp.4009–4018. 9 10

5. BorkmanS. CrespiA. DhakadS. GangulyS. HoginsJ. JhangY.-C. KamalzadehM. LiB. LealS. ParisiP. RomeroC. SmithW. ThamanA. WarrenS. YadavN.: Unity perception: Generate synthetic data for computer vision.CoRR abs/2107.04259(2021). 3

Cited by 1 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Data-driven Digital Lighting Design for Residential Indoor Spaces;ACM Transactions on Graphics;2023-03-17