Semantic Representation of Physical Activity Sensor Observations and Comparative Analysis of Real and Synthetic Datasets: A Proof-of-Concept-Study with MOX2-5 Sensor

Author:

Chatterjee Ayan1,Gerdes Martin W.1,Prinz Andreas1,Riegler Michael A.2,Martinez Santiago G.1

Affiliation:

1. University of Agder

2. Simula Metropolitan Center for Digital Engineering

Abstract

Abstract Background Daily activity of humans is monitored at a large scale automatically by devices such as mobile phones and wearables. This produces immense amounts of data that can be used to get a better understanding of human behavior over time. To understand this data and its possibilities, a structured and controlled collection process is required. Physical activity monitoring using wearable sensors has attracted prevalent attention in healthcare, sports science, and fitness applications. However, ensuring the availability of diverse and comprehensive datasets for research and algorithm development can be challenging. Objective We emphasize the importance of semantic representation for physical activity sensor observations to enable data interoperability and advanced analytics. In this proof-of-concept study, we propose an approach to improve the usability of physical activity datasets and highlight ethical considerations by generating synthetic datasets using medical-grade (CE certified) sensor. Moreover, our study presents a comparative analysis between real and synthetic activity datasets, evaluating their utilities to address model bias and fairness in predictive analysis. Methods We design and develop an ontology for semantic representation of physical activity sensor observations and predictive analysis on collected data with MOX2-5 activity sensors. The MOX2-5 activity monitoring device can collect and transmit high-resolution activity data such as activity intensity, weight-bearing, sedentary, standing, low physical activity, moderate physical activity, vigorous physical activity, and steps per minute. We collected physical activity data from 16 adults (Male: 12; Female: 4) for 30–45 days (about 1 and a half months). It produced a volume of 539 records which is small. Thus, we utilize different synthetic data generation methods, such as Gaussian Capula (GC), Conditional Tabular General Adversarial Network (CTGAN), and Tabular General Adversarial Network (TABGAN) to enhance the dataset with synthetic data. For both the real and synthetic datasets, we developed a Multilayer Perceptron (MLP) classification model to classify daily physical activity levels. Results The results highlight that semantic ontology is suitable for semantic search, knowledge representation, data integration, reasoning, and capturing the meaning and relationships between data. The analysis proves the hypothesis that the efficiency of predictive models grows with the increasing volume of additional synthetic training data. Conclusions The potential of ontology and Generative AI may accelerate research and innovation in the field of behavioral monitoring. Moreover, the presented data (both real MOX2-5 and its synthetic version) will be helpful in the creation of robust methods for the classification of activity types and different research directions in connection to synthetic data such as model efficiency, detection of generated data and data privacy.

Publisher

Research Square Platform LLC

Reference23 articles.

1. Benefits of Physical Activity. Webpage: https://www.cdc.gov/physicalactivity/basics/pa-health/index.htm. (Acceded on 18th September 2023).

2. ‘ProHealth eCoach: User-centered design and development of an eCoach app to promote healthy lifestyle with personalized activity recommendations’;Chatterjee A;BMC Health Services Research,2022

3. Physical activity. Webpage: https://www.who.int/news-room/fact-sheets/detail/physical-activity. (Acceded on 18th September 2023).

4. ‘Impact of activity monitoring on physical activity, sedentary behavior, and body weight during the COVID-19 pandemic’;Barkley JE;International Journal of Environmental Research and Public Health,2021

5. Thambawita, V. et al. (2020) ‘PMDATA’, Proceedings of the 11th ACM Multimedia Systems Conference [Preprint]. doi:10.1145/3339825.3394926.

同舟云学术

1.学者识别学者识别

2.学术分析学术分析

3.人才评估人才评估

"同舟云学术"是以全球学者为主线,采集、加工和组织学术论文而形成的新型学术文献查询和分析系统,可以对全球学者进行文献检索和人才价值评估。用户可以通过关注某些学科领域的顶尖人物而持续追踪该领域的学科进展和研究前沿。经过近期的数据扩容,当前同舟云学术共收录了国内外主流学术期刊6万余种,收集的期刊论文及会议论文总量共计约1.5亿篇,并以每天添加12000余篇中外论文的速度递增。我们也可以为用户提供个性化、定制化的学者数据。欢迎来电咨询!咨询电话:010-8811{复制后删除}0370

www.globalauthorid.com

TOP

Copyright © 2019-2024 北京同舟云网络信息技术有限公司
京公网安备11010802033243号  京ICP备18003416号-3