Let now count the fitted people, so has shape . Its rows have unit length, but they need not average to zero. Denote their mean row by . In the example it is
Let denote a column of ones, so repeats the mean row for every person. Subtract it from and call the resulting centered matrix :
Let denote person ’s row of . For A,
This second centering asks how A’s normalized coverage differs from the fitted population of normalized person profiles. The earlier subtraction compared articles inside a cell. The reference objects are different.
Let denote the people covariance. Using the fitted-person count , compute it as
Here and each index news coordinates from 1 to . Entry is the average product of those two centered coordinates across people. A large positive product tends to occur when both coordinates deviate in the same direction; a negative product occurs when they deviate oppositely. The diagonal entries are coordinate variances.
Each person contributes one row. A person supported by 10,000 articles does not enter the covariance a thousand times more heavily than a person supported by ten. More articles can affect the quality of a row, but they are not its weight in this fit. This is equal-person PCA: principal component analysis of normalized, centered coverage profiles.
Let index a people direction, its unit eigenvector column, and its covariance eigenvalue. Each pair satisfies
Multiplying the direction by the covariance stretches it by without turning it. As explained in the Eigen Times eigenvector chapter, the largest eigenvalue identifies the direction of greatest variation, and the next identifies the greatest variation perpendicular to it.
For the original four-person fixture, the first centered coordinate is approximately 0.3411 for A, 0.8207 for B, -0.3459 for C, and -0.8159 for D. The N1 variance therefore adds their four squares and divides by four, giving approximately 0.393847. The N1–N3 covariance instead multiplies each person’s first and third centered entries and averages the four products, giving approximately 0.138797. Both calculations use one product per person. The notebook prints the complete centered table and checks the covariance against independent reference entries.
Set the four-person matrix aside for a small covariance illustration. Let . The star marks this separate synthetic table. Let and be unit columns. Multiplying the first gives . Multiplying the second gives . They are eigenvectors with eigenvalues 3 and 1; their dot product is .
For a unit column , the scalar is variance along that direction. Along the first coordinate axis it is 2. Along it is 3; along it is 1. To see why 3 is the maximum, express any unit direction as , where are scalar coefficients satisfying . Its variance is , no larger than 3. The larger-eigenvalue direction captures the most spread because of this numerical property, not because its words are necessarily more important.
Multiplying an eigenvector by -1 leaves its line and eigenvalue unchanged. Its coordinates and resulting scores both reverse signs. A sign convention is useful for reproducible displays; it does not change the represented variation.
The toy covariance has eigenvalues approximately , , and . Keeping the first two retains
of the between-person variance, or 81.14%. This is a reconstruction statement about this four-person cloud. It is not an accuracy estimate, an identity confidence, or the fraction of articles explained.
Let count the retained positive people directions; it is 2 in this toy example. The bare is a count, distinct from the contrast row . The retained direction index runs from 1 to . Put those eigenvectors, ordered by decreasing eigenvalue, into the columns of the loading matrix :
Its rows correspond to N1, N2, N3; its columns correspond to people patterns P1 and P2. The columns have length one and are perpendicular. Their signs are oriented so the largest-magnitude news loading is positive, matching the implementation’s sign convention. Flipping a column and all its scores would describe the same geometry. When two eigenvalues are close, a small change in data can rotate their directions substantially even while their combined subspace changes little. Fixing signs does not remove that ambiguity; it is another reason to keep model versions explicit.
The three fitted eigenvalues sum to approximately 0.9596456. This is also the sum of the three diagonal entries of , called its trace: total centered squared length per person, expressed in either coordinate system. The first two sum to approximately 0.7786802. Their ratio is 0.8114248; multiplying by 100 expresses it as a percentage. The omitted fraction is about 0.1885752.
The loading columns are unit directions, but a row of need not have length one. Similarly, the score rows need not be unit rows after population centering and projection. Keeping these distinctions prevents a loading, a coordinate, and a cosine from being treated as interchangeable numbers.
Let denote person ’s people-coordinate row. Project its centered profile onto the retained columns:
For A, the first coordinate is approximately
The second uses the second column of . The resulting rows are:
| Person | P1 score | P2 score |
|---|---|---|
| A | 0.5109 | -0.5334 |
| B | 0.5828 | -0.1178 |
| C | 0.0813 | 0.8807 |
| D | -1.1750 | -0.2294 |
P1 is a pattern, not person D, although D has the largest absolute P1 score in this small example. A negative score identifies the other pole of a direction. It does not say that D is opposed to the other people.
In the published models, there are 60 news coordinates and at most 24 directions with positive eigenvalues. The 217-person Hacker News model retains 84.54% of between-person variance; the 61-person general-news model retains 90.86%. These separately fitted percentages should not be read as a contest between corpora. There is no second varimax naming step and no whitening of people scores. The people directions are orthonormal in standardized news coordinates, not generally when mapped back into the original embedding geometry.
A’s first score adds approximately . Its second coordinate is
These are projections of the same centered row onto different direction columns. Small differences in final digits arise if we multiply displayed four-decimal entries rather than the full-precision fixture.
The first score column has variance 0.4970 and the second 0.2817, the eigenvalues already reported. They are not divided by the square roots of those variances in this model. Dividing them would whiten the retained scores and would change distances and cosines; that is a different geometry.