The Consistency of Probabilistic Databases with Independent Cells

12/23/2022
by   Amir Gilad, et al.
0

A probabilistic database with attribute-level uncertainty consists of relations where cells of some attributes may hold probability distributions rather than deterministic content. Such databases arise, implicitly or explicitly, in the context of noisy operations such as missing data imputation, where we automatically fill in missing values, column prediction, where we predict unknown attributes, and database cleaning (and repairing), where we replace the original values due to detected errors or violation of integrity constraints. We study the computational complexity of problems that regard the selection of cell values in the presence of integrity constraints. More precisely, we focus on functional dependencies and study three problems: (1) deciding whether the constraints can be satisfied by any choice of values, (2) finding a most probable such choice, and (3) calculating the probability of satisfying the constraints. The data complexity of these problems is determined by the combination of the set of functional dependencies and the collection of uncertain attributes. We give full classifications into tractable and intractable complexities for several classes of constraints, including a single dependency, matching constraints, and unary functional dependencies.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
06/15/2022

On the complexity of finding set repairs for data-graphs

In the deeply interconnected world we live in, pieces of information lin...
research
05/04/2019

Learning Functional Dependencies with Sparse Regression

We study the problem of discovering functional dependencies (FD) from a ...
research
05/10/2022

Explainable Data Imputation using Constraints

Data values in a dataset can be missing or anomalous due to mishandling ...
research
03/26/2021

Synthesizing Linked Data Under Cardinality and Integrity Constraints

The generation of synthetic data is useful in multiple aspects, from tes...
research
09/29/2020

The Shapley Value of Inconsistency Measures for Functional Dependencies

Quantifying the inconsistency of a database is motivated by various goal...
research
09/29/2020

Database Repairing with Soft Functional Dependencies

A common interpretation of soft constraints penalizes the database for e...
research
08/19/2021

Temporal Graph Functional Dependencies: Technical Report

Data dependencies have been extended to graphs e.g., graph functional de...

Please sign up or login with your details

Forgot password? Click here to reset