Data Preparation and Exploration

Data Preparation and Exploration

Robert Hoyt / Robert Muenchen

33,44 €
IVA incluido
Disponible
Editorial:
Lulu Press
Año de edición:
2020
ISBN:
9780988752979
33,44 €
IVA incluido
Disponible

Selecciona una librería:

  • Librería Samer Atenea
  • Librería Aciertas (Toledo)
  • Kálamo Books
  • Librería Perelló (Valencia)
  • Librería Elías (Asturias)
  • Donde los libros
  • Librería Kolima (Madrid)
  • Librería Proteo (Málaga)

Data scientists spend more than two-thirds of their time cleaning, preparing, exploring, and visualizing data before it is ready for modeling and mining. This textbook covers the important steps of data preparation and exploration that anyone who deals with data should know. This textbook is an excellent companion text for our other textbook Introduction to Biomedical Data Science. The data preparation and exploration methods we include are spreadsheet and statistics package approaches, as well as the programming languages R and Python. The reader is introduced to the free stat packages Jamovi and BlueSky Statistics. Multiple techniques for data visualization are presented. Medical datasets are used for demonstrations and student exercises. Importantly, chapter content is supplemented with YouTube videos. Chapters are well referenced (100+) and there is a chapter on health data resources so the reader can find data to prepare and explore on their own.Prominent issues such as how to handle missing data and imbalanced datasets are covered along with sections on descriptive statistics, visualization, correlations, handling duplicates and outliers, scaling, standardization, and much more. A downloadable Data Checklist is available on https://www.informaticseducation.org

Artículos relacionados

  • Exploring Advances in Interdisciplinary Data Mining and Analytics
    Data mining is still a relatively young field, expanding at the rate of technology while advancing tools and techniques for gaining knowledge, finding patterns, and managing databases. Exploring Advances in Interdisciplinary Data Mining and Analytics: New Trends is an updated look at the state of technology in the field of data mining and analytics. As processor speeds, databas...
  • Knowledge Discovery Practices and Emerging Applications of Data Mining
    Recent developments have drastically increased the volume and complexity of data available to be mined, leading researchers to explore new ways to glean non-trivial data automatically. Knowledge Discovery Practices and Emerging Applications of Data Mining: Trends and New Domains introduces the reader to recent research activities in the field of data mining. This book covers as...
  • Research and Trends in Data Mining Technologies and Applications
    David Taniar
    ...
  • Developing Metadata Application Profiles
    The prevalence of data science has grown exponentially in recent years. Increases in data exchange have created the need for standards and formats on handling data from different sources. Developing Metadata Application Profiles is an innovative reference source that discusses the latest trends and techniques for effectively managing and exchanging metadata. Including a range o...
  • Modern Technologies for Big Data Classification and Clustering
    Data has increased due to the growing use of web applications and communication devices. It is necessary to develop new techniques of managing data in order to ensure adequate usage. Modern Technologies for Big Data Classification and Clustering is an essential reference source for the latest scholarly research on handling large data sets with conventional data mining and provi...
  • 90 Gelöste Fälle zu Zeitintelligenz in der DAX-Sprache
    Ramón Javier Castro Amador
    Dieser Ratgeber ist rein praktisch ausgerichtet, so dass Sie den gesamten DAX-Code in dieser Publikation anhand einer zum Download verfügbaren .pbix-Datei testen können.'90 gelöste Fälle zu Zeitintelligenz in DAX' ist ein Ratgeber für Benutzer von Microsoft Power BI, der Lösungen für sehr häufige praktische Fälle in Zeitintelligenzmodellen in der Sprache DAX bietet.Um das Verst...
    Disponible

    16,15 €

Otros libros del autor

  • Introduction to Biomedical Data Science
    Robert Hoyt / Robert Muenchen
    Introduction to Biomedical Data Science aims to fill the data science knowledge gap experienced by many clinical, administrative and technical staff. The textbook begins with an overview of what biomedical data science is and then embarks on a tour of topics beginning with spreadsheet tips and tricks and ending with artificial intelligence. In between, important topics are cove...
    Disponible

    79,19 €