Mining Imperfect Data

- With Examples in R and Python

Ronald K. Pearson

Bog

Format
Bog, paperback
Engelsk
481 sider

Indgår i serie
- Mathematics in Industry

Normalpris

kr. 899,95

Medlemspris

kr. 839,95

Du sparer kr. 60,00
Fri fragt

Som medlem af Saxo Premium 20 timer køber du til medlemspris, får fri fragt og 20 timers streaming/md. i Saxo-appen. De første 7 dage er gratis for nye medlemmer, derefter koster det 99,-/md. og kan altid opsiges. Løbende medlemskab, der forudsætter betaling med kreditkort. Fortrydelsesret i medfør af Forbrugeraftaleloven. Mindstepris 0 kr. Læs mere

Leveringstid: 8-11 Hverdage (Sendes fra fjernlager)
Forventet levering: 03-03-2026
Kan pakkes ind og sendes som gave
Split betalingen op med

Beskrivelse

It has been estimated that as much as 80% of the total effort in a typical data analysis project is taken up with data preparation, including reconciling and merging data from different sources, identifying and interpreting various data anomalies, and selecting and implementing appropriate treatment strategies for the anomalies that are found. This book focuses on the identification and treatment of data anomalies, including examples that highlight different types of anomalies, their potential consequences if left undetected and untreated, and options for dealing with them.

As both data sources and free, open-source data analysis software environments proliferate, more people and organizations are motivated to extract useful insights and information from data of many different kinds (e.g., numerical, categorical, and text). The book emphasizes the range of open-source tools available for identifying and treating data anomalies, mostly in R but also with several examples in Python.

Mining Imperfect Data: With Examples in R and Python, Second Editionpresents a unified coverage of 10 different types of data anomalies (outliers, missing data, inliers, metadata errors, misalignment errors, thin levels in categorical variables, noninformative variables, duplicated records, coarsening of numerical data, and target leakage);includes an in-depth treatment of time-series outliers and simple nonlinear digital filtering strategies for dealing with them; andprovides a detailed introduction to several useful mathematical characteristics of important data characterizations that do not appear to be widely known among practitioners, such as functional equations and key inequalities.

Læs hele beskrivelsen

Detaljer

SprogEngelsk
Sidetal481
Udgivelsesdato30-11-2020
ISBN139781611976267
Forlag Society For Industrial & Applied Mathematics,u.S.
FormatPaperback
UdgaveSecond Edition

Størrelse og vægt

Vægt1035 g

10 cm

Anmeldelser

Vær den første!

Log ind for at skrive en anmeldelse.

Mining Imperfect Data

- With Examples in R and Python

Beskrivelse

Anmeldelser

Mediernes anmeldelser

Findes i disse kategorier...