computers > database administration & management
Foundations of Data Quality Management
Imported Edition - Ships in 12-14 Days
Free Shipping in India on orders above Rs. 500
Imported Edition - Ships in 12-14 Days
Free Shipping in India on orders above Rs. 500
Data quality is one of the most important problems in data management. A database system typically aims to support the creation, maintenance, and use of large amount of data, focusing on the quantity of data. However, real-life data are often dirty: inconsistent, duplicated, inaccurate, incomplete, or stale. Dirty data in a database routinely generate misleading or biased analytical results and decisions, and lead to loss of revenues, credibility and customers. With this comes the need for data quality management. In contrast to traditional data management tasks, data quality management enables the detection and correction of errors in the data, syntactic or semantic, in order to improve the quality of the data and hence, add value to business processes. While data quality has been a longstanding problem for decades, the prevalent use of the Web has increased the risks, on an unprecedented scale, of creating and propagating dirty data. This monograph gives an overview of fundamental issues underlying central aspects of data quality, namely, data consistency, data deduplication, data accuracy, data currency, and information completeness. We promote a uniform logical framework for dealing with these issues, based on data quality rules. The text is organized into seven chapters, focusing on relational data. Chapter One introduces data quality issues. A conditional dependency theory is developed in Chapter Two, for capturing data inconsistencies. It is followed by practical techniques in Chapter 2b for discovering conditional dependencies, and for detecting inconsistencies and repairing data based on conditional dependencies. Matching dependencies are introduced in Chapter Three, as matching rules for data deduplication. A theory of relative information completeness is studied in Chapter Four, revising the classical Closed World Assumption and the Open World Assumption, to characterize incomplete information in the real world. A data currency model is presented in Chapter Five, to identify the current values of entities in a database and to answer queries with the current values, in the absence of reliable timestamps. Finally, interactions between these data quality issues are explored in Chapter Six. Important theoretical results and practical algorithms are covered, but formal proofs are omitted. The bibliographical notes contain pointers to papers in which the results were presented and proven, as well as references to materials for further reading. This text is intended for a seminar course at the graduate level. It is also to serve as a useful resource for researchers and practitioners who are interested in the study of data quality. The fundamental research on data quality draws on several areas, including mathematical logic, computational complexity and database theory. It has raised as many questions as it has answered, and is a rich source of questions and vitality. Table of Contents: Data Quality: An Overview / Conditional Dependencies / Cleaning Data with Conditional Dependencies / Data Deduplication / Information Completeness / Data Currency / Interactions between Data Quality Issues
Wenfei Fan is the (Chair) Professor of Web Data Management in the School of Informatics, University of Edinburgh, UK. He is a Fellow of the Royal Society of Edinburgh, UK, a National Professor of the 1000-Talent Program, and a Yangtze River Scholar, China. He received his Ph.D. from the University of Pennsylvania, U.S.A., and his M.S .and B.S. from Peking University, China. He is a recipient of the Alberto O. Mendelzon Test-of-Time Award of ACM PODS 2010, the Best Paper Award for VLDB 2010, the Roger Needham Award in 2008 (UK), the Best Paper Award for IEEE ICDE 2007, the Outstanding Overseas Young Scholar Award in 2003 (China), the Best Paper of the Year Award for Computer Networks in 2002, and the Career Award in 2001 (USA). His current research interests include database theory and systems, in particular data quality, data integration, database security, distributed query processing, query languages, social network analysis, Web services, and XML. Floris Geerts is Research Professorin the Department of Mathematics and Computer Science, University of Antwerp, Belgium. Before that, he held a senior research fellow position in the database group at the University of Edinburgh, UK and a postdoctoral research position in the data mining group at the University of Helsinki, Finland. He received his Ph.D. in 2001 from the University of Hasselt, Belgium. His research interests include the theory and practice of databases and the study of data quality, in particular. He is a recipient of the Best Paper Awards for IEEE ICDM 2001 and IEEE ICDE 2007.
• Author(s): Clear | James • Publisher: Penguin • Publisher Imprint: Penguin Random House • Subject: General Books
• Author(s): Jeff Kinney • Publisher: Penguin Random House Children's UK • Publisher Imprint: Penguin Random House Children's UK • BISAC: Comics & Graphic Novels - Humorous
• Author(s): Ichiro Kishimi • Publisher: GROVE ATLANTIC • Publisher Imprint: Allen & Unwin • BISAC: Personal Growth - SuccessIchiro Kishimi lives in Kyoto. He writes, lectures and teaches in psychiatric clinics as a certified counsellor and c...
View full details• Author(s): Chetan Bhagat • Publisher: HarperCollins Publishers India • Publisher Imprint: HarperCollins Publishers India • BISAC: GeneralFrom India's top-selling writer Chetan Bhagat comes a powerful new love story that will make you laugh, cry...
View full details• Author(s): Brianna Wiest • Publisher: Manjul Publishing • Publisher Imprint: Amaryllis • BISAC: Body Mind And SpiritThis is a book about self-sabotage. Why we do it, when we do it, and how to stop doing it—for good. Coexisting but conflicting n...
View full details• Author(s): Morgan Housel • Publisher: Pan Macmillan • Publisher Imprint: Pan Macmillan • BISAC: Finance - Wealth ManagementA third book from the International bestselling author of The Psychology of Money and Same as Ever, lessons on harnessing...
View full details• Author(s): Arundhati Roy• Publisher: PRH INDIA LOCAL PRINT• Publisher Imprint: Penguin Hamish Hamilton• BISAC: Literary FiguresArundhati Roy’s first work of memoir, this is a soaring account, both intimate and inspiring, of how the author became...
View full details• Author(s): Acharya Prashant • Publisher: HarperCollins Publishers India • Publisher Imprint: HarperCollins Publishers India • BISAC: GeneralIn a world where vagueness is mistaken for depth and obscurity passes for wisdom, Truth without Apology ...
View full details• Author(s): Sudha Murthy • Publisher: India Puffin • Publisher Imprint: India Puffin • BISAC: Short StoriesWho can resist a good story, especially when it's being told by Grandma? From her bag emerges tales of kings and cheats, monkeys and mic...
View full details• Author(s): Satoshi Yagisawa • Publisher: Bonnier Books Ltd • Publisher Imprint: Bonnier Books Ltd
• Author(s): Newport, Cal • Publisher: Little, Brown Book Group • Publisher Imprint: Piatkus
• Author(s): Shrijeet Shandilya • Publisher: Ebury Press • Publisher Imprint: Ebury Press • BISAC: Romance - GeneralIn the electric haze of college life, three friends are bound by laughter, late-night talks and unspoken promises. But when two of...
View full details• Author(s): Dan Brown • Publisher: Transworld Publishers Ltd • Publisher Imprint: Transworld Publishers Ltd • BISAC: Thrillers - EspionageDan Brown is the bestselling author of Digital Fortress, Deception Point, Angels and Demons, The Da Vinci C...
View full details• Author(s): Sudha Murty • Publisher: India Puffin • Publisher Imprint: India Puffin • BISAC: Action & Adventure - General
Rich Dad Poor Dad: What the Rich Teach Their Kids about Money That the Poor and Middle Class Do Not!
• Publisher: Penguin • Publisher Imprint: Penguin Random House • Subject: General Books • BISAC: Personal Finance - GeneralApril of 2022 marks a 25-year milestone for the personal finance classic Rich Dad Poor Dad that still ranks as the #1 Pers...
View full details• Author(s): Dale Carnegie | Napoleon Hill • Publisher: Fingerprint • Publisher Imprint: Fingerprint • Subject: General Books
• Author(s): Freida Mcfadden • Publisher: Penguin Select Print • Publisher Imprint: Penguin Select Publishing"Multi-Million Copy Bestselling Series •Now Being Made Into a Major Motion Picture Starring Sydney Sweeney and Amanda Seyfried #1 New Yor...
View full details• Author(s): Wonder House Books • Publisher: Wonder House Books • Publisher Imprint: Wonder House Books • BISAC: Comics & Graphic Novels - Fairy Tales, Folklore, Legends & MTimeless Wisdom, Talking Animals & Life Lessons for Young Min...
View full details• Author(s): Viktor E. Frankl • Publisher: Random House • Publisher Imprint: Random Hou • Subject: Medical, Nursing and Health Sciences
• Author(s): Madhavi Bharadwaj • Publisher: PRH India • Publisher Imprint: Penguin Ebury Press • BISAC: Parenting - MotherhoodWelcome to the wild, messy, wonderful world of parenting--where the nights are long, the diapers are explosive, and unso...
View full details• Author(s): Vir Das • Publisher: HarperCollins Publishers India • Publisher Imprint: HarperCollins Publishers India • BISAC: Entertainment & Performing ArtsComedian and actor Vir Das is beloved (by some, tolerated by others, blocked by a few...
View full details• Author(s): Eric Carle • Publisher: Penguin Books, Limited (UK) • Publisher Imprint: Penguin Books, Limited (UK) • BISAC: Animals - Butterflies, Moths & CaterpillarsEric Carle's The Very Hungry Caterpillar is a perennial favourite with child...
View full details• Author(s): Prajakta Koli • Publisher: Harper Fiction India • Publisher Imprint: Harper Fiction India • BISAC: Romance - ContemporaryWinner of the Amazon India Popular Choice Debut Book 2025 Award. From one of India's most-loved creators comes s...
View full details