The quantification of gesture–speech synchrony: A tutorial and validation of multimodal data acquisition using device-based and video-based motion tracking-Reference-Cited by-同舟云学术

The quantification of gesture–speech synchrony: A tutorial and validation of multimodal data acquisition using device-based and video-based motion tracking

Published:2019-10-28 Issue:2 Volume:52 Page:723-740
ISSN:1554-3528
Container-title:Behavior Research Methods
language:en
Short-container-title:Behav Res

Author:

Pouw Wim^ORCID,Trujillo James P.,Dixon James A.

Abstract

Abstract There is increasing evidence that hand gestures and speech synchronize their activity on multiple dimensions and timescales. For example, gesture’s kinematic peaks (e.g., maximum speed) are coupled with prosodic markers in speech. Such coupling operates on very short timescales at the level of syllables (200 ms), and therefore requires high-resolution measurement of gesture kinematics and speech acoustics. High-resolution speech analysis is common for gesture studies, given that field’s classic ties with (psycho)linguistics. However, the field has lagged behind in the objective study of gesture kinematics (e.g., as compared to research on instrumental action). Often kinematic peaks in gesture are measured by eye, where a “moment of maximum effort” is determined by several raters. In the present article, we provide a tutorial on more efficient methods to quantify the temporal properties of gesture kinematics, in which we focus on common challenges and possible solutions that come with the complexities of studying multimodal language. We further introduce and compare, using an actual gesture dataset (392 gesture events), the performance of two video-based motion-tracking methods (deep learning vs. pixel change) against a high-performance wired motion-tracking system (Polhemus Liberty). We show that the videography methods perform well in the temporal estimation of kinematic peaks, and thus provide a cheap alternative to expensive motion-tracking systems. We hope that the present article incites gesture researchers to embark on the widespread objective study of gesture kinematics and their relation to speech.

Funder

The Netherlands Organisation of Scientific Research

Publisher

Springer Science and Business Media LLC

Subject

General Psychology,Psychology (miscellaneous),Arts and Humanities (miscellaneous),Developmental and Educational Psychology,Experimental and Cognitive Psychology

Link

http://link.springer.com/content/pdf/10.3758/s13428-019-01271-9.pdf

Reference63 articles.

1. Alexanderson, S., House, D., & Beskow, J. (2013, August). Aspects of co-occurring syllables and head nods in spontaneous dialogue. Paper presented at the 12th International Conference on Auditory–Visual Speech Processing (AVSP 2013), Annecy, France.

2. Alviar, C., Dale, R., & Galati, A. (2019). Complex communication dynamics: Exploring the structure of an academic talk. Cognitive Science, 43, e12718. https://doi.org/10.1111/cogs.12718

3. Anzulewicz, A., Sobota, K., & Delafield-Butt, J. T. (2016). Toward the autism motor signature: Gesture patterns during smart tablet gameplay identify children with autism. Scientific reports, 6, 31107.

4. Beckman, M. E., & Ayers, G. (1997). Guidelines for ToBI labelling, version 3. The Ohio State University Research Foundation. Retrieved from http://www.ling.ohio-state.edU/phonetics/ToBI/ToBI.0.html .

5. Beecks, C., Hassani, M., Hinnell, J., Schüller, D., Brenger, B., Mittelberg, I., & Seidl, T. (2015). Spatiotemporal similarity search in 3d motion capture gesture streams. In International Symposium on Spatial and Temporal Databases (pp. 355–372). Cham, Switzerland: Springer.

Cited by 42 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Charting the Silent Signals of Social Gaze: Automating Eye Contact Assessment in Face-to-Face Conversations;2024-08-29

2. Anaphoric demonstratives occur with fewer and different pointing gestures than exophoric demonstratives;Glossa: a journal of general linguistics;2024-08-27

3. Enhancing Assessment of Social Motor Synchrony Through Full-Body Interaction: A Novel Approach with OSMoSIS Tool;Proceedings of the 23rd Annual ACM Interaction Design and Children Conference;2024-06-17

4. Phonetic differences between affirmative and feedback head nods in German Sign Language (DGS): A pose estimation study;PLOS ONE;2024-05-30

5. How and Why People Synchronize: An Integrated Perspective;Personality and Social Psychology Review;2024-05-21