
Interactively Exploring High-Dimensional Data and Models in R by Dianne Cook
Visualizing data is a powerful tool for uncovering patterns and insights that might otherwise remain hidden. While there are numerous resources available for data visualization, few focus comprehensively on high-dimensional data visualization. High-dimensional data, or multivariate data, arises when multiple variables are measured for each observation, presenting unique challenges and opportunities for analysis. High-dimensional data visualisation is valuable for understanding dimension reduction methods, unsupervised and supervised classification. This book provides a detailed guide to visualizing high-dimensional data and models using linear projections, with practical examples and R code to help readers explore these fascinating data spaces.
Through this book, readers will learn how to identify patterns, clusters, and anomalies in high-dimensional data that are often obscured in lower-dimensional plots. By integrating visualization techniques with analytical methods, the book aims to enhance the understanding and interpretation of complex data structures, making it an essential resource for anyone working with multivariate data. The book is organised into three parts, following overview and introductory chapters. The dimension reduction chapters cover principal component analysis and nonlinear dimension reduction. The chapters on cluster analysis cover hierarchical and k-means algorithms, model-based and self-organising maps, and finish with ways to communicate results and how to compare different results. The chapters on classification cover linear discriminant analysis, tree and forest algorithms, support vector machines and neural networks.
Key Features
- Comprehensive Introduction: Learn the fundamentals of high-dimensional spaces, visualization techniques, and essential notation for advanced methods.
- Dimension Reduction Techniques: Explore linear and non-linear methods to summarize high-dimensional data, detect issues, and evaluate representation quality.
- Cluster Analysis: Discover graphical and numerical approaches to identify groups in data, assess clustering techniques, and visualize solutions in high dimensions.
- Classification Methods: Understand how to explore known groups, check model assumptions, examine classification boundaries, and identify errors.
- Integration with R: Includes R code examples using packages like tourr, detourr, and mulgar to complement explanations and plots.
- Toolbox Chapter: A dedicated appendix chapter provides an overview of primary visualization methods and guidance for getting started.
This book is designed for students, educators, researchers, data analysts, and industry professionals working in fields such as biology, social sciences, finance, and machine learning. It is particularly suited for those engaged in exploratory data analysis and model fitting for multivariate data. To make effective use of this material the reader should have a basic working knowledge of R and some understanding of multivariate statistical methods or machine learning methods.
-
Advanced R
-
Using R for Introductory Statistics
-
R Markdown
-
Statistical Computing with R, Second Edition
-
Hands-On Machine Learning with R
-
Graphical Data Analysis with R
-
Introduction to Scientific Programming and Simulation Using R
-
Advanced R, Second Edition
-
Extending R
-
Reproducible Research with R and R Studio
-
Work Automation with R
- R for Social Network Analysis
-
Stated Preference Methods Using R
-
Introduction to Forestry Data Analysis with R
-
Simulation and Power Analysis Using R
-
R and MATLAB
-
The Essentials of Data Science: Knowledge Discovery Using R
-
blogdown
-
Statistical Computing in C++ and R
-
Model-Based Clustering, Classification, and Density Estimation Using mclust in R
-
Using R for Modelling and Quantitative Methods in Fisheries
-
Analyzing Baseball Data with R
-
Parallel Computing for Data Science
-
Computational Actuarial Science with R
-
Using the R Commander
-
Implementing Reproducible Research
-
R for Conservation and Development Projects
-
Distributions for Modeling Location, Scale, and Shape
-
Interactive Web-Based Data Visualization with R, plotly, and shiny
-
R Markdown Cookbook
-
Learn R
-
Displaying Time Series, Spatial, and Space-Time Data with R
-
Copula Additive Distributional Regression Using R
-
Introductory Fisheries Analyses with R
-
Introduction to Political Analysis in R
-
Engineering Production-Grade Shiny Apps
Dianne Cook and Ursula Laa have jointly published numerous papers on methodology for high-dimensional data visualisation in the past decade. This book is a result of these collaborations. Dianne Cook has been researching methods for data visualisation, particularly for exploratory data analysis, and data mining, for more than 30 years. She is a Distinguished Professor of Statistics at Monash University, Fellow of the American Statistical Association, past editor of the Journal of Computational and Graphical Statistics, and the R Journal, Board Member of the R Foundation, and elected member of the International Statistical Institute, and author of numerous R packages. Ursula Laa is an Assistant Professor at the Institute of Statistics of the University of Natural Resources and Life Sciences in Vienna. She works on new methods for the visualisation of multivariate data and models, and on interdisciplinary applications of statistics and data science methods in different fields.
| SKU | Unavailable |
| ISBN 13 | 9781032748443 |
| ISBN 10 | 1032748443 |
| Title | Interactively Exploring High-Dimensional Data and Models in R |
| Author | Dianne Cook |
| Series | Chapman And Hall Crc The R Series |
| Condition | Unavailable |
| Binding Type | Hardback |
| Publisher | Taylor & Francis Ltd |
| Year published | 2025-10-10 |
| Number of pages | 272 |
| Cover note | Book picture is for illustrative purposes only, actual binding, cover or edition may vary. |
| Note | Unavailable |


































