Supervised versus unsupervised approaches to classification of accelerometry data-Reference-Cited by-同舟云学术

Supervised versus unsupervised approaches to classification of accelerometry data

Published:2023-05 Issue:5 Volume:13 Page:
ISSN:2045-7758
Container-title:Ecology and Evolution
language:en
Short-container-title:Ecology and Evolution

Author:

Sur Maitreyi¹,Hall Jonathan C.²,Brandt Joseph³,Astell Molly³⁴,Poessel Sharon A.⁵^ORCID,Katzner Todd E.⁵^ORCID

Affiliation:

1. Conservation Science Global, Inc. West Cape May New Jersey USA

2. Department of Biology Eastern Michigan University Ypsilanti Michigan USA

3. U.S. Fish and Wildlife Service, Hopper Mountain National Wildlife Refuge Complex Ventura California USA

4. Department of Biology Boise State University Boise Idaho USA

5. U.S. Geological Survey, Forest and Rangeland Ecosystem Science Center Boise Idaho USA

Abstract

AbstractSophisticated animal‐borne sensor systems are increasingly providing novel insight into how animals behave and move. Despite their widespread use in ecology, the diversity and expanding quality and quantity of data they produce have created a need for robust analytical methods for biological interpretation. Machine learning tools are often used to meet this need. However, their relative effectiveness is not well known and, in the case of unsupervised tools, given that they do not use validation data, their accuracy can be difficult to assess. We evaluated the effectiveness of supervised (n = 6), semi‐supervised (n = 1), and unsupervised (n = 2) approaches to analyzing accelerometry data collected from critically endangered California condors (Gymnogyps californianus). Unsupervised K‐means and EM (expectation–maximization) clustering approaches performed poorly, with adequate classification accuracies of <0.8 but very low values for kappa statistics (range: −0.02 to 0.06). The semi‐supervised nearest mean classifier was moderately effective at classification, with an overall classification accuracy of 0.61 but effective classification only of two of the four behavioral classes. Supervised random forest (RF) and k‐nearest neighbor (kNN) machine learning models were most effective at classification across all behavior types, with overall accuracies >0.81. Kappa statistics were also highest for RF and kNN, in most cases substantially greater than for other modeling approaches. Unsupervised modeling, which is commonly used for the classification of a priori‐defined behaviors in telemetry data, can provide useful information but likely is instead better suited to post hoc definition of generalized behavioral states. This work also shows the potential for substantial variation in classification accuracy among different machine learning approaches and among different metrics of accuracy. As such, when analyzing biotelemetry data, best practices appear to call for the evaluation of several machine learning techniques and several measures of accuracy for each dataset under consideration.

Funder

California Department of Fish and Wildlife

National Fish and Wildlife Foundation

Publisher

Wiley

Subject

Nature and Landscape Conservation,Ecology,Ecology, Evolution, Behavior and Systematics

Link

https://onlinelibrary.wiley.com/doi/pdf/10.1002/ece3.10035

Reference60 articles.

1. Classifying behavior from short‐interval biologging data: An example with GPS tracking of birds

2. The roller coaster flight strategy of bar-headed geese conserves energy during Himalayan migrations

3. Optimizing acceleration-based ethograms: the use of variable-time versus fixed-time segmentation

4. Observing the unwatchable through acceleration logging of animal behavior

Cited by 2 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Thanks to repetition, dustbathing detection can be automated combining accelerometry and wavelet analysis;Ethology;2024-04-24

2. How to account for behavioral states in step-selection analysis: a model comparison;PeerJ;2024-02-26