How Google Actually Evaluates Links: PageRank-NearestSeeds and Why Zero-Traffic Links Are Worthless

Off-Page SEO Guide

How Google Actually Evaluates Links: PageRank-NearestSeeds and Why Zero-Traffic Links Are Worthless

In May 2024, roughly 2,500 pages of Google’s internal search documentation leaked. It didn’t hand us the algorithm, but it did confirm how links are scored: through a version of PageRank measured from trusted seed sites, weighted by the quality and traffic of the page the link sits on. Here’s what that means for the links you chase.

By Rahul Saini, Author at Search Counsel Co. Last updated [July] 2026.

Featured answer: how does Google evaluate a backlink?

Google scores a link using PageRank-NearestSeeds, a version of PageRank that measures how far the linking page sits from a set of trusted “seed” sites. The closer and more trusted the source, the more value passes. A link’s worth also rises with the quality, freshness, and traffic of the page it’s on, and its topical relevance to yours. A link from a page with no clicks and no trust passes little or nothing.

The Source

~2,500 pages

of internal Google API docs leaked in May 2024, later confirmed authentic.

The Metric

PageRank-NS

named in the docs as the production PageRank value, measured from seed sites.

The Catch

No clicks, no value

links on low-tier pages that earn no clicks appear to pass little or no PageRank.

The Proof

Links still move

rankings dropped when Ahrefs disavowed links, then recovered when it undid it.

Most “what makes a good backlink” advice is guesswork dressed up as fact, because Google keeps its ranking systems private. The 2024 documentation leak is the closest thing we’ve had to a look under the hood. It doesn’t reveal weightings or prove any feature is live in ranking, and Google was quick to caution against over-reading it. But it names the link features Google’s systems track, and those names line up with what SEOs have measured for years. This guide walks through how Google appears to evaluate a link: PageRank and seed sites, the tier and traffic of the source page, homepage trust, relevance and anchors, and the spam controls. Then it turns all of that into what you should actually do.

Article note: Written by Rahul Saini at Search Counsel Co. This guide reads the leaked Google Content Warehouse API documentation and the analyses of it by Rand Fishkin, Mike King, and Search Engine Land, alongside Ahrefs’ own link experiments. One firm caveat throughout: the leak describes attributes Google’s systems store, not confirmed ranking rules or weightings. Google verified the documents were real but warned against assuming how, or whether, any feature is used. Treat everything below as strongly-informed interpretation, and treat DR and DA as third-party estimates, not Google metrics.

1) What the leak actually was

In late May 2024, about 2,500 pages of internal documentation for Google’s Content Warehouse API surfaced publicly, first surfaced by Rand Fishkin and Mike King. Google later confirmed the documents were authentic. What they contain is a catalog of “features,” the data attributes Google’s systems can store about a page, a link, or a site, with short internal descriptions. What they don’t contain is the recipe: no weightings, no confirmation of what’s used in live ranking. So the leak is a vocabulary list, not an answer key. That still matters, because for the first time we can see which link signals Google bothers to track, and the list quietly contradicts years of official “we don’t use that” statements. For the confirmed side of the picture, see our guide to Google ranking factors.

Read this first: a feature existing in the docs doesn’t prove Google ranks with it. But Google doesn’t build and maintain a data attribute for no reason. When the leaked features line up with independent ranking studies, the signal gets a lot more credible. That’s the standard used throughout this guide.

2) PageRank-NearestSeeds and seed sites

PageRank never died. It evolved. The leaked docs reference a feature called pageRankNS, described as “PageRank-NearestSeeds,” and note it as the production PageRank value teams should use, while the original PageRank is flagged as long-deprecated. That single line reframes link building, and it’s a useful companion to how search engines work more generally.

Here’s the concept, and it isn’t complicated. Instead of calculating authority across the whole messy web, Google appears to start from a set of trusted “seed” sites and measure how far every other page sits from them. The closer your linking page is to a seed, in link hops, the more value it can pass. Google has only ever publicly named two seeds, The New York Times and the now-defunct Google Directory, and the concept traces back to a Google patent analyzed at length by the late Bill Slawski. In practice, seed sites are a topic’s most authoritative, well-connected publications. A link from one of them, or from a page one hop away, is worth far more than a link from a site buried 40 hops out in the link graph.

  • Proximity is the point. A relevant link from a trusted publication beats a pile of links from small, isolated sites that no seed site would ever link to.
  • This is why digital PR wins. Earning a mention in a real publication puts you close to a seed. Directory submissions and forum drops do not. See our guide to digital PR for link building.
  • Small isn’t disqualifying. A tiny but topically perfect site can still pass value if it’s itself close to trusted sources. Proximity plus relevance, not size alone.

3) Source quality, tier and freshness

The docs include a feature called sourceType that records the quality of a link’s source page, tied to the index tier the page lives in. Analysts read the tiers as roughly high, medium, and low quality (TYPE_HIGH_QUALITY, TYPE_MEDIUM_QUALITY, TYPE_LOW_QUALITY), and the description notes it “can be used as an importance indicator” for the link. Plainly put: a link from a page in Google’s top tier is worth more than the same link from a low-tier page.

Freshness gets special treatment. The docs mark freshly published content (TYPE_FRESHDOCS) as equivalent to high quality for the purpose of link importance. That’s a real strategic nudge: a link from a brand-new, high-quality article can carry more weight than one on a page that’s sat untouched for a decade. It also explains why an ongoing stream of new links tends to outperform a one-time burst that then goes stale, and why a content refresh cycle helps on the receiving end too.

4) Why zero-traffic links are worthless

This is the finding that changes how you vet a link. Per the analysis of the leak, Google appears to use clicks to a page to decide which quality tier that page belongs in. A page that earns real clicks lands in a higher tier; a page that gets none tends to fall into the low-quality tier. And links from low-tier pages appear to pass little or no PageRank. In other words, a link on a page nobody visits may do nothing for you at all.

Be careful with certainty here, because this is interpretation, not a confirmed rule, and correlation isn’t causation. But it lines up with hard evidence from two directions. First, Ahrefs’ 2025 study of links from pages with traffic found the strongest single correlate of a link’s value was the referring page’s own link strength (URL Rating), which tracks closely with whether a page is alive and well-linked. Second, when Ahrefs disavowed the links to three of its own pages, rankings and traffic fell, then recovered once the disavow was lifted, confirming that real links carry real weight. Put together, the practical rule is blunt: a link is only as good as the page it lives on, and a dead page is a dead link. Our guide to what makes a good backlink turns that into a five-minute check.

From experience: before we value any link opportunity, we check whether the target page actually gets traffic and ranks for anything. A guest post slot on a “DR 60” site means nothing if the article will sit on a page with zero visits. We’d take an in-content mention on a smaller page that genuinely gets read over that every time. This is the single most common way people waste a link budget, and it’s worth remembering when evaluating guest posting offers.

5) Homepage trust and site authority

The leak shows Google tracking PageRank and trust at the homepage level (homepagePagerankNs and homepage-trust signals), and separately a feature literally named siteAuthority inside a quality-signals module. Two takeaways follow. First, site-wide authority is a real thing Google measures, even if it isn’t computed the way Moz’s DA or Ahrefs’ DR are, so those third-party scores are rough proxies at best. Second, trust appears to flow from the homepage inward: strengthening your homepage’s authority can lift the inner pages it links to, which is a strong argument for deliberate internal linking and a clean site architecture. It’s also why a link to your homepage from a trusted source has value beyond the page it points at, it raises the trust that trickles across your whole site.

6) Relevance, anchors and mismatch

Relevance shows up on both sides of the link. The docs reference topical weighting on links and a demotion for anchor mismatch, where the anchor text and the destination page don’t line up. That confirms two long-held rules. A link from a topically related page is worth more, and your anchor text should honestly describe where the link goes. It also flags a risk: stuffing your profile with exact-match keyword anchors is a pattern Google can detect and discount, and at scale it correlates with penalties. Natural, varied, descriptive anchors are the goal. We cover the safe distribution in anchor text best practices.

One more useful detail from the docs: there’s no “dofollow” attribute, only the absence of a nofollow tag, and the documentation suggests that if a page links to you more than once and any single instance is followed, the link can be treated as followed. It’s a small point, but it’s another reason not to obsess over the follow tag on any single mention. For the full picture of what separates a strong link from a weak one, start with what makes a good backlink.

7) How Google neutralizes spam links

The docs are full of spam controls, which tells you Google expects manipulation and plans for it. Features reference link velocity (acquiring many links too fast can be flagged), anchor-spam measures, and Penguin timestamps. The takeaway isn’t fear. It’s that most junk links are simply ignored rather than punished, so a few bad links won’t sink you. The danger is a deliberate pattern: a sudden spike of paid links with matching exact-match anchors, or a profile dominated by low-tier spam, the sort of tactics covered in white hat vs black hat SEO. Google’s spam systems (Penguin, and the machine-learning SpamBrain) are built to discount those automatically. That’s also why the disavow tool is a last resort, not routine hygiene, a point we cover in toxic backlinks and the disavow tool.

Leaked feature What it appears to track What it means for you
pageRankNS PageRank measured from trusted seed sites (the production value). Earn links close to authoritative publications, not from isolated sites.
sourceType Quality tier of the linking page, with fresh pages treated as high quality. Prioritize links on strong, current pages; keep earning new ones.
Clicks / tier signals Pages with no clicks appear to drop to the low-quality tier. A link on a page with no traffic likely passes little or nothing.
homepagePagerankNs / siteAuthority Homepage trust and a site-wide authority score. Build homepage authority; trust flows to inner pages.
Topicality / anchor mismatch Relevance of the link and alignment of anchor to destination. Get relevant links; use natural, descriptive anchors.
Link velocity / anchor spam Unnatural link spikes and over-optimized anchors. Grow links steadily; avoid exact-match anchors at scale.

8) What to do about all this

Strip away the feature names and the leak points to one strategy, the same one good SEOs already followed, now with evidence behind it.

  1. Chase proximity, not volume. A few links close to trusted seed sites beat hundreds from the far edges of the web. That means digital PR and genuine editorial mentions.
  2. Only value links on living pages. Before you pursue a link, check the target page gets traffic and ranks for something. Skip dead pages, however high the domain score.
  3. Keep earning fresh links. A steady stream of new links from current, high-quality pages outperforms a one-time burst. Freshness is a real advantage.
  4. Build homepage trust. Links to your homepage from respected sources lift the whole site, not just one page.
  5. Stay relevant and natural. Topically related links, descriptive anchors, no exact-match stuffing, no unnatural spikes.

The single highest-leverage move is to become the kind of source seed sites cite: original data, real research, genuinely useful tools. That earns links close to the seeds and, increasingly, citations in AI answers built on the same trust signals. We break that down in original research that earns links and citations, in off-page SEO for AI, and in our guide to getting cited in AI search.

How we do it: at Search Counsel Co. we build links the way Google’s own systems reward, close to trusted sources, on pages that get read, relevant and natural, earned rather than bought. If you’d rather hand it off, see our link building and authority services, or start with the full off-page SEO and link building guide.

9) Sources used for this guide

This is a data-led piece, so it leans on primary analyses of the leak and on named link experiments rather than opinion.

Source What it supports
Google Content Warehouse API leak (May 2024) The feature names: pageRankNS, sourceType, homepage PageRank, siteAuthority, anchor mismatch, link velocity.
Rand Fishkin (SparkToro) and Mike King (iPullRank) Initial breakdowns, including the clicks-to-tier interpretation and the deprecation of original PageRank.
Search Engine Land / Digitaloft (link-builder takeaways) PageRank-NearestSeeds as the production value; sourceType tiers and freshdocs; seed-site concept.
Bill Slawski (SEO by the Sea) patent analysis The seed-site model and how distance from seeds affects scoring.
Ahrefs (links-with-traffic study, 2025; disavow / impact-of-links test) URL Rating as the strongest link correlate; rankings drop when links are disavowed and recover when restored.
Google public statement on the leak Confirmation the docs are authentic, with a caution against assuming feature usage or weighting.

FAQ: how Google evaluates links

Does Google still use PageRank?

Yes, in an evolved form. The leaked documentation names PageRank-NearestSeeds as the production PageRank value and flags the original 1998 PageRank as long-deprecated. So the concept of links as votes is very much alive; it’s just measured relative to trusted seed sites now, not across the raw web.

What are seed sites?

Seed sites are a set of highly trusted, well-connected websites Google appears to use as reference points, measuring how far every other page sits from them. The closer a linking page is to a seed, the more value it can pass. Google has only publicly named The New York Times and the old Google Directory, but seeds are effectively a topic’s most authoritative publications.

Are links from pages with no traffic worthless?

Largely, yes, based on the leak analysis. Google appears to use clicks to sort pages into quality tiers, and links from low-tier pages seem to pass little or no PageRank. It isn’t a confirmed rule, but it matches independent data, so the safe move is to only value links on pages that get real traffic.

Is Domain Authority (DA) or Domain Rating (DR) a Google ranking factor?

No. The leak shows Google has its own site-authority signal, but it isn’t computed like Moz’s DA or Ahrefs’ DR. Those are third-party estimates. They’re handy for quick comparisons and screening, but don’t mistake them for how Google actually scores your site.

Did the API leak confirm these are ranking factors?

Not exactly. The leak confirms Google stores these attributes, and Google confirmed the documents are real. But Google cautioned that you can’t assume how, or whether, a feature is used in live ranking. The features are strong evidence, especially where they match ranking studies, not a rulebook.

How does this change my link building?

Favor relevant links close to trusted publications, on pages that get real traffic and are freshly published, with natural anchors, earned steadily over time. That’s digital PR and original data, not directories, forum drops, or bought links on dead pages.

Conclusion: build the links Google’s own systems reward

The 2024 leak didn’t give us the algorithm, but it settled a lot of arguments. Links are scored by proximity to trusted seed sites, by the quality, freshness, and traffic of the page they sit on, and by relevance. A link on a page nobody reads is close to worthless, and no domain score changes that. So stop counting links and start earning the few that sit near the sources Google already trusts.

For the practical playbook, read how to get backlinks that actually work, and to judge the links you already have, run a backlink audit.

Editorial note: This guide interprets leaked internal documentation, not confirmed Google ranking rules. Feature descriptions, third-party authority scores, and correlation studies are indicators, not guarantees. Verify current details against primary sources before relying on them, and remember correlation does not prove causation.

 

Scroll to Top