Affiliation:
1. School of Water Conservancy and Transportation, Zhengzhou University, Zhengzhou 450001, China
2. CEC Guiyang Exploration and Design Research Institute Co., Guiyang 550081, China
Abstract
The development of information technology has led to massive, multidimensional, and heterogeneously sourced disaster data. However, there’s currently no universal metadata standard for managing natural disasters. Common pre-training models for information extraction requiring extensive training data show somewhat limited effectiveness, with limited annotated resources. This study establishes a unified natural disaster metadata standard, utilizes self-trained universal information extraction (UIE) models and Python libraries to extract metadata stored in both structured and unstructured forms, and analyzes the results using the Word2vec-Kmeans cluster algorithm. The results show that (1) the self-trained UIE model, with a learning rate of 3 × 10−4 and a batch_size of 32, significantly improves extraction results for various natural disasters by over 50%. Our optimized UIE model outperforms many other extraction methods in terms of precision, recall, and F1 scores. (2) The quality assessments of consistency, completeness, and accuracy for ten tables all exceed 0.80, with variances between the three dimensions being 0.04, 0.03, and 0.05. The overall evaluation of data items of tables also exceeds 0.80, consistent with the results at the table level. The metadata model framework constructed in this study demonstrates high-quality stability. (3) Taking the flood dataset as an example, clustering reveals five main themes with high similarity within clusters, and the differences between clusters are deemed significant relative to the differences within clusters at a significance level of 0.01. Overall, this experiment supports effective sharing of disaster data resources and enhances natural disaster emergency response efficiency.
Funder
National Key Research and Development Program of China
Henan provincial key research and development program
Reference53 articles.
1. Application of Social Sensors in Natural Disasters Emergency Management: A Review;Shi;IEEE Trans. Comput. Soc. Syst.,2023
2. Parallelizing Word2Vec in Shared and Distributed Memory;Ji;IEEE Trans. Parallel Distrib. Syst.,2019
3. Method of Multi-type Disaster Data Organization and Management Based on GeoSOT;Liao;Geogr. Geo-Inf. Sci.,2013
4. Jony, R.I., Woodley, A., and Perrin, D. (2019, January 2–4). Flood Detection in Social Media Images using Visual Features and Metadata. Proceedings of the 2019 Digital Image Computing: Techniques and Applications (DICTA), Perth, WA, Australia.
5. Tian, Y., and Li, W. (2022). GeoAI for Knowledge Graph Construction: Identifying Causality Between Cascading Events to Support Environmental Resilience Research arXiv. arXiv.