The Consistency of Probabilistic Databases with Independent Cells

Amir Gilad*, Aviram Imber*, Benny Kimelfeld*

*Corresponding author for this work

Research output: Chapter in Book/Report/Conference proceedingConference contributionpeer-review

1 Scopus citations

Abstract

A probabilistic database with attribute-level uncertainty consists of relations where cells of some attributes may hold probability distributions rather than deterministic content. Such databases arise, implicitly or explicitly, in the context of noisy operations such as missing data imputation, where we automatically fill in missing values, column prediction, where we predict unknown attributes, and database cleaning (and repairing), where we replace the original values due to detected errors or violation of integrity constraints. We study the computational complexity of problems that regard the selection of cell values in the presence of integrity constraints. More precisely, we focus on functional dependencies and study three problems: (1) deciding whether the constraints can be satisfied by any choice of values, (2) finding a most probable such choice, and (3) calculating the probability of satisfying the constraints. The data complexity of these problems is determined by the combination of the set of functional dependencies and the collection of uncertain attributes. We give full classifications into tractable and intractable complexities for several classes of constraints, including a single dependency, matching constraints, and unary functional dependencies.

Original languageEnglish
Title of host publication26th International Conference on Database Theory, ICDT 2023
EditorsFloris Geerts, Brecht Vandevoort
PublisherSchloss Dagstuhl- Leibniz-Zentrum fur Informatik GmbH, Dagstuhl Publishing
ISBN (Electronic)9783959772709
DOIs
StatePublished - 1 Mar 2023
Externally publishedYes
Event26th International Conference on Database Theory, ICDT 2023 - Ioannina, Greece
Duration: 28 Mar 202331 Mar 2023

Publication series

NameLeibniz International Proceedings in Informatics, LIPIcs
Volume255
ISSN (Print)1868-8969

Conference

Conference26th International Conference on Database Theory, ICDT 2023
Country/TerritoryGreece
CityIoannina
Period28/03/2331/03/23

Bibliographical note

Publisher Copyright:
© Amir Gilad, Aviram Imber, and Benny Kimelfeld; licensed under Creative Commons License CC-BY 4.0 26th International Conference on Database Theory (ICDT 2023)

Keywords

  • attribute-level uncertainty
  • functional dependencies
  • most probable database
  • Probabilistic databases

Fingerprint

Dive into the research topics of 'The Consistency of Probabilistic Databases with Independent Cells'. Together they form a unique fingerprint.

Cite this