Supporting User-Defined Functions on Uncertain Data

Adler R. J.; Antova L.; Bishop C. M.; Chaudhuri S.; Dalvi N. N.; Denny M.; Deshpande A.; Gibbs A. L.; Girard A.; Kurose J. F.; McLachlan G.; Nguyen D. T.; O'Hagan A.; Ranganathan A.; Rasmussen C. E.; Sen P.; Singh S.; Szalay A. S.; Tran T.; Tran T. T. L.

research

Supporting User-Defined Functions on Uncertain Data

Authors: Adler R. J.
Antova L.
Bishop C. M.
Chaudhuri S.
Dalvi N. N.
Denny M.
Deshpande A.
Gibbs A. L.
Girard A.
Kurose J. F.
McLachlan G.
Nguyen D. T.
O'Hagan A.
Ranganathan A.
Rasmussen C. E.
Sen P.
Singh S.
Szalay A. S.
Tran T.
Tran T. T. L.
Publication date: 1 January 2013
Publisher
Doi

Abstract

Uncertain data management has become crucial in many sensing and scientific applications. As user-defined functions (UDFs) become widely used in these applications, an important task is to capture result uncertainty for queries that evaluate UDFs on uncertain data. In this work, we provide a general framework for supporting UDFs on uncertain data. Specifically, we propose a learning approach based on Gaussian processes (GPs) to compute approximate output distributions of a UDF when evaluated on uncertain input, with guaranteed error bounds. We also devise an online algorithm to compute such output distributions, which employs a suite of optimizations to improve accuracy and performance. Our evaluation using both real-world and synthetic functions shows that our proposed GP approach can outperform the state-of-the-art sampling approach with up to two orders of magnitude improvement for a variety of UDFs. 1

Similar works

Full text

Open in the Core reader

Download PDF

Available Versions

Edinburgh Research Explorer

oai:pure.ed.ac.uk:publications...

Last time updated on 08/02/2015

Crossref

info:doi/10.14778%2F2536336.25...

Last time updated on 01/04/2019