Eigenvalues, diagonalisation and the spectral theorem
LA · Chapter 412 min readAsked at Two Sigma, Citadel, DE Shaw, Jane Street
After this lesson you should be able to
- Say what an eigenvector is and why diagonalisation is useful.
- State the spectral theorem and what it guarantees for symmetric matrices.
- Connect the largest eigenvalue to power iteration and to Markov chains.
An eigenvector is a direction the matrix does not rotate — it only stretches. Finding those directions turns a complicated linear map into independent one-dimensional scalings, which is why eigendecomposition is the tool behind PCA, Markov chain limits, and every question about repeated application of a matrix.
Equation 4.1
The eigenvalue equation
A direction that is merely scaled, and the characteristic polynomial whose roots are the scalings.
- The stretch factor. Zero means the direction is collapsed.
- Defined up to scale, so only the direction is determined.
Proposition 4.2
Diagonalisation and why it helps
When , applying repeatedly becomes trivial: , and raising a diagonal matrix to a power means raising each entry. Anything involving repeated application — a Markov chain, a difference equation, a matrix exponential — collapses to independent scalar problems in the eigenbasis.
Holds when
- Not every matrix is diagonalisable; repeated eigenvalues can leave it defective.
- Every *symmetric* matrix is, which is why the theory is so much cleaner in that case.
- Similar matrices share eigenvalues, so a change of basis does not change the spectrum.
Definition 4.3
The spectral theorem
Spectral theorem, — A real symmetric matrix has real eigenvalues and an orthonormal basis of eigenvectors. Since covariance matrices, projection matrices and Hessians are all symmetric, this covers nearly everything that turns up in statistics — and it is the theorem that makes PCA work.
Why symmetry buys orthogonality. A symmetric matrix acts the same way in both directions: . Apply that to two eigenvectors with different eigenvalues and you get , which forces the inner product to zero. Orthogonality is not an extra assumption but a two-line consequence — and it is what lets PCA produce uncorrelated components rather than merely a change of basis.
Example 4.4
Find the eigenvalues of the correlation matrix with off-diagonal , and say what they mean.
Show the worked solutionHide the worked solution
Worked solution
- Formula
- Substitute
- Solve
- Answer
Sanity check. They sum to 2, the trace, as they must. The first eigenvector is — the common move — and carries of the variance; the second is , the spread between them, with . That is a two-asset PCA done by hand.
The rest of this lesson is in Premium
You have read the opening. 10 more sections follow, including 4 worked examples and 3 quick checks.
Nothing is charged for 7 days, and you can cancel before then. Or read The law of large numbers and the central limit theorem in full, free.