| tr(S), digit “8” | 741.158872 | 1001 |
| λ1 | 151.568412 (20.45% of the total) | 1002 |
| uncentred b1 vs the mean image | 0.2851∘ | 1001 |
| centred vs uncentred b1 | 87.9996∘ | 1001 |
| rank of the centred data matrix | 52 of 64; 12 dead pixels | 1001 |
| storage break-even | M=47, i.e. ND/(N+D)=46.79 | 1001 |
| bm⊤Sbm vs λm | 8.5×10−14 | 1002 |
| 200,000 random unit vectors vs λ1 | 37.18% | 1002 |
| components for 50/90/99% of the variance | 5 / 18 / 36 | 1002 |
| measured JM vs ∑j>Mλj | 2.8×10−13 over all M | 1002, 1004 |
| Eq 10.32 vs a 100,001-point grid search | 4.919350 vs 4.919360 | 1003 |
| the Pythagoras identity, 200,000 random z | 1.7×10−15 | 1003 |
| non-orthonormal basis, same subspace | 24.728409 vs 16.316272 | 1003 |
| share of the error in U⊥ | 1.000000000000 | 1003 |
| VM+JM−tr(S) | 2.3×10−13 | 1004 |
| 200,000 random orthonormal bases, M=5 | best 624.426067 vs 320.123369; 0 beat it | 1004 |
| rotating B: JM, codes | unchanged; 39.243537 | 1004 |
| code covariance off-diagonal, eigenbasis vs rotated | 5.1×10−14 vs 37.046852 | 1004 |
| λd vs σd2/N | 1.4×10−13 | 1005 |
| κ(S)/κ(X) | 246.8358=κ(X) | 1005 |
eigh at κ=1010 | 5.3×103 relative, 18 negatives | 1005 |
| Eckart–Young spectral error vs σM+1 | 7×10−14 | 1005 |
| power-iteration rate, predicted vs observed | 0.578708 vs 0.578541 | 1005 |
| steps to 10−6 deg at λ2/λ1=0.99 | 1769, against 27 at 0.5 | 1005 |
| N=50 in D=784: nonzero eigenvalues | 49 | 1006 |
| Gram matrix condition number (centred) | 4.15×1017 | 1006 |
| N×N speed-up at D=4000 | 4565.2× | 1006 |
| unnormalised Xcm | 6.269×104 vs 1.160×101 | 1006 |
| metres vs millimetres, unstandardised | 77.3982∘; standardised, 1.2×10−6 | 1007 |
| tr, raw vs standardised | 386.429897 vs 2.000000=D | 1007 |
| NaNs from Equation 10.58 | 2,088 | 1007 |
| test statistics vs training statistics | 15.400250 vs 15.781992 | 1007 |
| Eq 10.70b by 400,000 samples | 2.1×10−3 relative (floor 1.6×10−3) | 1008 |
| C over 200 observations | 0.000×100 change | 1008 |
| Cmm vs σ2/λm | off-diagonal 9.7×10−16 | 1008 |
| σML2 vs JM/(D−M) | identical to 8 decimals | 1008 |
| PPCA shrinkage at σ2=40 | 0.736093 down to 0.236834 | 1008 |
| M=0 baseline, RMS | 27.224233 | 1009 |
| four rules for M | 1, 18, 25, 38 | 1009 |
| rotation vs rescaling equivariance | 8.0×10−16 vs 59.2214∘ | 1009 |
| PPCA imputation vs mean-filling | 36.8% better at 10% missing | 1009 |
| novelty: digit “0” flagged | 100.0% | 1009 |
| a noisy circle’s eigenvalues | 0.501038, 0.497606 | 1009 |
| kernel PCA on two rings | separation 0.0597→7.5299 | 1009 |