Normal view MARC view ISBD view

Data profiling

By: Abedjan, Ziawasch.
Contributor(s): Golab, Lukasz | Naumann, Felix | Papenbrock, Thorsten.
Material type: materialTypeLabelBookPublisher: S.l. : Morgan & Claypool Publisher , 2019Description: xviii, 136 p. : ill. ; 23.5 cm.ISBN: 9781681734460 .Subject(s): Metadata | Data miningDDC classification: 025.3 Summary: Data profiling refers to the activity of collecting data about data, i.e., metadata. Most IT professionals and researchers who work with data have engaged in data profiling, at least informally, to understand and explore an unfamiliar dataset or to determine whether a new dataset is appropriate for a particular task at hand. Data profiling results are also important in a variety of other situations, including query optimization, data integration, and data cleaning. Simple metadata are statistics, such as the number of rows and columns, schema and datatype information, the number of distinct values, statistical value distributions, and the number of null or empty values in each column. More complex types of metadata are statements about multiple columns and their correlation, such as candidate keys, functional dependencies, and other types of dependencies.
Tags from this library: No tags from this library for this title. Log in to add tags.
Item type Current location Call number Status Date due Barcode
Books 025.3 ABE (Browse shelf) Available 031831

Data profiling refers to the activity of collecting data about data, i.e., metadata. Most IT professionals and researchers who work with data have engaged in data profiling, at least informally, to understand and explore an unfamiliar dataset or to determine whether a new dataset is appropriate for a particular task at hand. Data profiling results are also important in a variety of other situations, including query optimization, data integration, and data cleaning. Simple metadata are statistics, such as the number of rows and columns, schema and datatype information, the number of distinct values, statistical value distributions, and the number of null or empty values in each column. More complex types of metadata are statements about multiple columns and their correlation, such as candidate keys, functional dependencies, and other types of dependencies.

There are no comments for this item.

Log in to your account to post a comment.

Powered by Koha