Verwandte Artikel zu Data Profiling (Synthesis Lectures on Data Management)

Data Profiling (Synthesis Lectures on Data Management) - Hardcover

 
9781681734484: Data Profiling (Synthesis Lectures on Data Management)

Zu dieser ISBN ist aktuell kein Angebot verfügbar.

Reseña del editor

Data profiling refers to the activity of collecting data about data, i.e., metadata. Most IT professionals and researchers who work with data have engaged in data profiling, at least informally, to understand and explore an unfamiliar dataset or to determine whether a new dataset is appropriate for a particular task at hand. Data profiling results are also important in a variety of other situations, including query optimization, data integration, and data cleaning. Simple metadata are statistics, such as the number of rows and columns, schema and datatype information, the number of distinct values, statistical value distributions, and the number of null or empty values in each column. More complex types of metadata are statements about multiple columns and their correlation, such as candidate keys, functional dependencies, and other types of dependencies.

This book provides a classification of the various types of profilable metadata, discusses popular data profiling tasks, and surveys state-of-the-art profiling algorithms. While most of the book focuses on tasks and algorithms for relational data profiling, we also briefly discuss systems and techniques for profiling non-relational data such as graphs and text. We conclude with a discussion of data profiling challenges and directions for future work in this area.

Biografía del autor

Ziawasch Abedjan is Juniorprofessor (Assistant Professor) and Head of the "Big Data Management" (BigDaMa) Group at the Technische Universität Berlin. Before Ziawasch was a postdoc at the "Computer Science and Artificial Intelligence Laboratory" at MIT working on various data integration topics. Ziawasch received his Ph.D. from the Hasso Plattner Institute in Potsdam, Germany. His research interests include, data mining, data integration, and data profiling.

„Über diesen Titel“ kann sich auf eine andere Ausgabe dieses Titels beziehen.

  • VerlagMorgan & Claypool Publishers
  • Erscheinungsdatum2018
  • ISBN 10 1681734486
  • ISBN 13 9781681734484
  • EinbandTapa dura
  • SpracheEnglisch
  • Anzahl der Seiten156

(Keine Angebote verfügbar)

Buch Finden:



Kaufgesuch aufgeben

Sie finden Ihr gewünschtes Buch nicht? Wir suchen weiter für Sie. Sobald einer unserer Buchverkäufer das Buch bei AbeBooks anbietet, werden wir Sie informieren!

Kaufgesuch aufgeben

Weitere beliebte Ausgaben desselben Titels

9781681734460: Data Profiling (Synthesis Lectures on Data Management)

Vorgestellte Ausgabe

ISBN 10:  168173446X ISBN 13:  9781681734460
Verlag: Morgan & Claypool Publishers, 2018
Softcover