Download now Free registration required
Data quality is a critical problem in modern databases. Data entry forms present the first and arguably best opportunity for detecting and mitigating errors, but there has been little research into automatic methods for improving data quality at entry time. In this paper, the authors propose Usher, an end-to-end system for form design, filling, and data quality assurance. Using previous form submissions, Usher learns a probabilistic model over the questions of the form. Usher then applies this model at every step of the data entry process to ensure high quality. Before entry, it induces a form layout that captures the most important data values of a form instance as quickly as possible.
- Format: PDF
- Size: 710.4 KB