A Classification Paradigm for Distributed Vertically Partitioned Data-Reference-Cited by-同舟云学术

A Classification Paradigm for Distributed Vertically Partitioned Data

Published:2004-07-01 Issue:7 Volume:16 Page:1525-1544
ISSN:0899-7667
Container-title:Neural Computation
language:en
Short-container-title:Neural Computation

Author:

Basak Jayanta¹,Kothari Ravi¹

Affiliation:

1. IBM India Research Laboratory, Indian Institute of Technology, New Delhi 110016, India

Abstract

In general, pattern classification algorithms assume that all the features are available during the construction of a classifier and its subsequent use. In many practical situations, data are recorded in different servers that are geographically apart, and each server observes features of local interest. The underlying infrastructure and other logistics (such as access control) in many cases do not permit continual synchronization. Each server thus has a partial view of the data in the sense that feature subsets (not necessarily disjoint) are available at each server. In this article, we present a classification algorithm for this distributed vertically partitioned data. We assume that local classifiers can be constructed based on the local partial views of the data available at each server. These local classifiers can be any one of the many standard classifiers (e.g., neuralnetworks, decision tree, k nearest neighbor). Often these local classifiers are constructed to support decision making at each location, and our focus is not on these individual local classifiers. Rather, our focus is constructing a classifier that can use these local classifiers to achieve an error rate that is as close as possible to that of a classifier having access to the entire feature set. We empirically demonstrate the efficacy of the proposed algorithm and also provide theoretical results quantifying the loss that results as compared to the situation where the entire feature set is available to any single classifier.

Publisher

MIT Press - Journals

Subject

Cognitive Neuroscience,Arts and Humanities (miscellaneous)

Link

https://www.mitpressjournals.org/doi/pdf/10.1162/089976604323057470

Reference6 articles.

1. Universal approximation bounds for superpositions of a sigmoidal function

2. Bagging predictors

3. Complexity measures of supervised classification problems

4. Adaptive Mixtures of Local Experts

5. Statistical pattern recognition: a review

Cited by 16 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Novel method for optimizing performance in resource constrained distributed data streams;Applied Intelligence;2022-02-16

2. A comparative evaluation of aggregation methods for machine learning over vertically partitioned data;Expert Systems with Applications;2020-08

3. HDSM: A distributed data mining approach to classifying vertically distributed data streams;Knowledge-Based Systems;2020-02

4. BIBLIOGRAPHY;Data Mining;2019-10-17

5. Supervised Classification Algorithms in Machine Learning: A Survey and Review;Advances in Intelligent Systems and Computing;2019-07-17