A PLS kernel algorithm for data sets with many variables and few objects. Part II: Cross-validation, missing data and examples
- Authors
- Stefan Rännar, Paul L. M. Geladi, Fredrik Lindgren and Svante Wold
- Published
- 1995
- DOI
- 10.1002/cem.1180090604
- Citation
- Rännar et al.: "A PLS kernel algorithm for data sets with many variables and few objects. Part II: Cross-validation, missing data and examples", Journal of Chemometrics, 9, 459-470, 1995.
- Abstract
- This is Part II of a series concerning the PLS kernel algorithm for data sets with many variables and few objects. Here the issues of cross-validation and missing data are investigated. Both partial and full crossvalidation are evaluated in terms of predictive residuals and speed and are illustrated on real examples. Two related approaches to the solution of the missing data problem are presented. One is a full EM algorithm and the second a reduced EM algorithm which applies when the number of missing values is small. The two examples are multivariate calibration data sets. The first set consists of UV–visible data measured on mixtures of four metal ions. The second example consists of FT-IR measurements on mixtures consisting of four different organic substances.
- Tags
Related items
- Principles of proper validation: use and abuse of re-sampling for validation (2010)
- Kernel-PCA algorithms for wide data Part II: fast cross-validation and application in classification of NIR data (1997)
- A PLS kernel algorithm for data sets with many variables and fewer objects. Part 1: Theory and algorithm (1994)
- The kernel PCA algorithms for wide data. Part I: theory and algorithms (1997)
- The treatment of missing measurements in PCA and PLS models (2002)