Kernel Principal Component Analysis - Large Datasets

Large Datasets

In practice, a large data set leads to a large K, and storing K may become a problem. One way to deal with this is to perform clustering on your large dataset, and populate the kernel with the means of those clusters. Since even this method may yield a relatively large K, it is common to compute only the top P eigenvalues and eigenvectors of K.

Read more about this topic:  Kernel Principal Component Analysis

Famous quotes containing the word large:

    Switzerland is a small, steep country, much more up and down than sideways, and is all stuck over with large brown hotels built on the cuckoo clock style of architecture.
    Ernest Hemingway (1899–1961)