Edinburgh Research Explorer

Towards Certain Fixes with Editing Rules and Master Data

Research output: Contribution to journalArticle

Related Edinburgh Organisations

Open Access permissions

Open

Original languageEnglish
Pages (from-to)173-184
Number of pages12
JournalProceedings of the VLDB Endowment (PVLDB)
Volume3
Issue number1
Publication statusPublished - 2010

Abstract

A variety of integrity constraints have been studied for data cleaning. While these constraints can detect the presence of errors, they fall short of guiding us to correct the errors. Indeed, data repairing based on these constraints may not
find certain xes that are absolutely correct, and worse, may introduce new errors when repairing the data. We propose a method for finding certain fixes, based on master data, a notion of certain regions, and a class of editing rules. A
certain region is a set of attributes that are assured correct by the users. Given a certain region and master data, editing rules tell us what attributes to fix and how to update them. We show how the method can be used in data monitoring
and enrichment. We develop techniques for reasoning about editing rules, to decide whether they lead to a unique fix and whether they are able to fix all the attributes in a tuple, relative to master data and a certain region. We also provide an algorithm to identify minimal certain regions, such that a certain fix is warranted by editing rules and master data as long as one of the regions is correct. We experimentally verify the effectiveness and scalability of the algorithm.

Download statistics

No data available

ID: 17653130