Get Free Shipping on orders over $79
Data Profiling : Synthesis Lectures on Data Management - Felix Naumann

Data Profiling

By: Felix Naumann, Ziawasch Abedjan, Thorsten Papenbrock, Lukasz Golab

Paperback | 8 November 2018

At a Glance

Paperback


$84.99

or 4 interest-free payments of $21.25 with

 or 

Ships in 7 to 10 business days

Data profiling refers to the activity of collecting data about data, {i.e.}, metadata. Most IT professionals and researchers who work with data have engaged in data profiling, at least informally, to understand and explore an unfamiliar dataset or to determine whether a new dataset is appropriate for a particular task at hand. Data profiling results are also important in a variety of other situations, including query optimization, data integration, and data cleaning. Simple metadata are statistics, such as the number of rows and columns, schema and datatype information, the number of distinct values, statistical value distributions, and the number of null or empty values in each column. More complex types of metadata are statements about multiple columns and their correlation, such as candidate keys, functional dependencies, and other types of dependencies.

This book provides a classification of the various types of profilable metadata, discusses popular data profiling tasks, and surveys state-of-the-art profiling algorithms. While most of the book focuses on tasks and algorithms for relational data profiling, we also briefly discuss systems and techniques for profiling non-relational data such as graphs and text. We conclude with a discussion of data profiling challenges and directions for future work in this area.

More in Internet Guides & Online Services

This Is for Everyone - Tim Berners-Lee

RRP $36.99

$29.75

20%
OFF
E-commerce 2023-2024 : 18 Edition - business. technology. society - Carol Traver
Customer Relationship Management : 2nd edition - Ed Peelen

RRP $157.45

$119.75

24%
OFF
Become a YouTuber : Build Your Own YouTube Channel - Cristina Calabrese
Digital Degrowth : Radically Rethinking our Digital Futures - Neil Selwyn
Blockchain : Blueprint for a New Economy - Melanie Swa

RRP $66.75

$30.99

54%
OFF
In This Economy? : How Money and Markets Really Work - Kyla Scanlan
MoneyGPT : AI and the Threat to the Global Economy - James Rickards
The Art of SEO : Mastering Search Engine Optimization - Eric Enge