View in EDS

Finding the K Highest-Ranked Answers in a Distributed Network

Saved in:

Bibliographic Details
Title:	Finding the K Highest-Ranked Answers in a Distributed Network
Authors:	Demetrios Zeinalipour-Yazti, Zografoula Vagena
Contributors:	The Pennsylvania State University CiteSeerX Archives
Source:	http://www.cs.ucr.edu/%7Etsotras/functional/24.pdf.
Publication Year:	2009
Collection:	CiteSeerX
Subject Terms:	Key words, Distributed Top-K Query Processing, P2P Networks, Sensor Networks
Description:	In this paper we present an algorithm for finding the k highest-ranked (or Top-k) answers in a distributed network. A Top-K query returns the subset of most relevant answers, in place of all answers, for two reasons: i) to minimize the cost metric that is associated with the retrieval of all answers; and ii) to improve the recall and the precision of the answer-set, such that the user is not overwhelmed with irrelevant results. Our study focuses on multi-hop distributed networks in which the data is accessible by traversing a network of nodes. Such a setting captures very well the computation framework of emerging Sensor Networks, Peer-to-Peer Networks and Vehicular Networks. We present the Threshold Join Algorithm (TJA), an efficient algorithm that utilizes a non-uniform threshold on the queried attribute in order to minimize the transfer of data when a query is executed. Additionally, TJA resolves queries in the network rather than in a centralized fashion which further minimizes the consumption of bandwidth and delay. We performed an extensive experimental evaluation of our algorithm using a real testbed of 75 workstations along with a trace-driven experimental methodology. Our results indicate that TJA requires an order of magnitude less communication than the state-of-the-art, scales well with respect to the parameter k and the network topology.
Document Type:	text
File Description:	application/pdf
Language:	English
Relation:	http://citeseerx.ist.psu.edu/viewdoc/summary?doi=10.1.1.190.9859; http://www.cs.ucr.edu/%7Etsotras/functional/24.pdf
Availability:	http://citeseerx.ist.psu.edu/viewdoc/summary?doi=10.1.1.190.9859 http://www.cs.ucr.edu/%7Etsotras/functional/24.pdf
Rights:	Metadata may be used without restrictions as long as the oai identifier remains attached to it.
Accession Number:	edsbas.81C6FF78
Database:	BASE

View record from BASE

Nájsť tento článok vo Web of Science

Description
Abstract:	In this paper we present an algorithm for finding the k highest-ranked (or Top-k) answers in a distributed network. A Top-K query returns the subset of most relevant answers, in place of all answers, for two reasons: i) to minimize the cost metric that is associated with the retrieval of all answers; and ii) to improve the recall and the precision of the answer-set, such that the user is not overwhelmed with irrelevant results. Our study focuses on multi-hop distributed networks in which the data is accessible by traversing a network of nodes. Such a setting captures very well the computation framework of emerging Sensor Networks, Peer-to-Peer Networks and Vehicular Networks. We present the Threshold Join Algorithm (TJA), an efficient algorithm that utilizes a non-uniform threshold on the queried attribute in order to minimize the transfer of data when a query is executed. Additionally, TJA resolves queries in the network rather than in a centralized fashion which further minimizes the consumption of bandwidth and delay. We performed an extensive experimental evaluation of our algorithm using a real testbed of 75 workstations along with a trace-driven experimental methodology. Our results indicate that TJA requires an order of magnitude less communication than the state-of-the-art, scales well with respect to the parameter k and the network topology.