A Randomized Algorithm for CCA

11/13/2014
by   Paul Mineiro, et al.
0

We present RandomizedCCA, a randomized algorithm for computing canonical analysis, suitable for large datasets stored either out of core or on a distributed file system. Accurate results can be obtained in as few as two data passes, which is relevant for distributed processing frameworks in which iteration is expensive (e.g., Hadoop). The strategy also provides an excellent initializer for standard iterative solutions.

READ FULL TEXT

Please sign up or login with your details

Forgot password? Click here to reset