0
Your cart

Your cart is empty

Books > Computing & IT > Computer communications & networking

Buy Now

Data Profiling (Paperback) Loot Price: R1,560
Discovery Miles 15 600
Data Profiling (Paperback): Ziawasch Abedjan, Lukasz Golab, Felix Naumann, Thorsten Papenbrock

Data Profiling (Paperback)

Ziawasch Abedjan, Lukasz Golab, Felix Naumann, Thorsten Papenbrock

Series: Synthesis Lectures on Data Management

 (sign in to rate)
Loot Price R1,560 Discovery Miles 15 600 | Repayment Terms: R146 pm x 12*

Bookmark and Share

Expected to ship within 10 - 15 working days

Data profiling refers to the activity of collecting data about data, {i.e.}, metadata. Most IT professionals and researchers who work with data have engaged in data profiling, at least informally, to understand and explore an unfamiliar dataset or to determine whether a new dataset is appropriate for a particular task at hand. Data profiling results are also important in a variety of other situations, including query optimization, data integration, and data cleaning. Simple metadata are statistics, such as the number of rows and columns, schema and datatype information, the number of distinct values, statistical value distributions, and the number of null or empty values in each column. More complex types of metadata are statements about multiple columns and their correlation, such as candidate keys, functional dependencies, and other types of dependencies. This book provides a classification of the various types of profilable metadata, discusses popular data profiling tasks, and surveys state-of-the-art profiling algorithms. While most of the book focuses on tasks and algorithms for relational data profiling, we also briefly discuss systems and techniques for profiling non-relational data such as graphs and text. We conclude with a discussion of data profiling challenges and directions for future work in this area.

General

Imprint: Springer International Publishing AG
Country of origin: Switzerland
Series: Synthesis Lectures on Data Management
Release date: November 2018
First published: 2019
Authors: Ziawasch Abedjan • Lukasz Golab • Felix Naumann • Thorsten Papenbrock
Dimensions: 235 x 191mm (L x W)
Format: Paperback
Pages: 136
ISBN-13: 978-3-03-100737-8
Languages: English
Subtitles: English
Categories: Books > Computing & IT > General theory of computing > Data structures
Books > Computing & IT > Computer programming > Algorithms & procedures
Books > Computing & IT > Computer communications & networking > General
LSN: 3-03-100737-9
Barcode: 9783031007378

Is the information for this product incomplete, wrong or inappropriate? Let us know about it.

Does this product have an incorrect or missing image? Send us a new image.

Is this product missing categories? Add more categories.

Review This Product

No reviews yet - be the first to create one!

Partners