Generating Activity Snippets by Learning Human-Scene Interactions-Reference-Cited by-同舟云学术

Generating Activity Snippets by Learning Human-Scene Interactions

Published:2023-07-26 Issue:4 Volume:42 Page:1-15
ISSN:0730-0301
Container-title:ACM Transactions on Graphics
language:en
Short-container-title:ACM Trans. Graph.

Author:

Li Changyang¹^ORCID,Yu Lap-Fai¹^ORCID

Affiliation:

1. George Mason University, Fairfax, United States of America

Abstract

We present an approach to generate virtual activity snippets, which comprise sequenced keyframes of multi-character, multi-object interaction scenarios in 3D environments, by learning from recordings of human-scene interactions. The generation consists of two stages. First, we use a sequential deep graph generative model with a temporal module to iteratively generate keyframe descriptions, which represent abstract interactions using graphs, while preserving spatial-temporal relations through the activities. Second, we devise an optimization framework to instantiate the activity snippets in virtual 3D environments guided by the generated keyframe descriptions. Our approach optimizes the poses of character and object instances encoded by the graph nodes to satisfy the relations and constraints encoded by the graph edges. The instantiation process includes a coarse 2D optimization followed by a fine 3D optimization to effectively explore the complex solution space for placing and posing the instances. Through experiments and a perceptual study, we applied our approach to generate plausible activity snippets under different settings.

Publisher

Association for Computing Machinery (ACM)

Subject

Computer Graphics and Computer-Aided Design

Link

https://dl.acm.org/doi/pdf/10.1145/3592096

Reference61 articles.

1. Text2SceneVR

2. Task-based locomotion

3. Nikos Athanasiou , Mathis Petrovich , Michael J Black , and Gül Varol . 2022 . TEACH: Temporal Action Composition for 3D Humans. arXiv preprint arXiv:2209.04066 (2022). Nikos Athanasiou, Mathis Petrovich, Michael J Black, and Gül Varol. 2022. TEACH: Temporal Action Composition for 3D Humans. arXiv preprint arXiv:2209.04066 (2022).

4. Synthesis of concurrent object manipulation tasks;Bai Yunfei;ACM Transactions on Graphics,2012

5. ActivityNet: A large-scale video benchmark for human activity understanding

Cited by 5 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Neural natural language processing for long texts: A survey on classification and summarization;Engineering Applications of Artificial Intelligence;2024-07

2. Utilizing passage‐level relevance and kernel pooling for enhancing BERT‐based document reranking;Computational Intelligence;2024-06

3. Experience Graph using Spatio-Temporal Scene Data for Replaying Mixed Reality Interaction;2024 IEEE Conference on Virtual Reality and 3D User Interfaces Abstracts and Workshops (VRW);2024-03-16

4. Simulating real-life scenarios to better understand the spread of diseases under different contexts;Scientific Reports;2024-02-01

5. Data Augmentation for Sample Efficient and Robust Document Ranking;ACM Transactions on Information Systems;2023-11-29