The Toolset

PageRank Crawl & Link Flow

Most tools checked a single page's score. The PageRank crawl went a step further: it fetched a page, found every link leaving it, and reported the authority of each destination. It was a window into how importance travelled across the web — the flow of what SEOs came to call "link equity." That flow is still central to how sites rank.

A glowing network of linked pages with particles flowing along the links, representing a link crawl and authority flow

What the crawl revealed

By listing the PageRank of every page a site linked out to, the crawl exposed a site's neighbourhood. Was a page linking to strong, reputable resources, or to a swamp of low-quality sites? Where was it sending its visitors — and, just as importantly, its authority? For anyone auditing a website, seeing the whole outbound picture at once was far more informative than any single score.

How authority flows through links

The original PageRank math treated a page's importance as something it passes along to the pages it links to. Picture it as water: importance pours into a page through its incoming links and then flows out, divided among that page's outgoing links. A link from a high-authority page delivers a bigger share than one from a weak page. This is the mechanism behind the endlessly repeated SEO phrase "link juice."

One practical consequence: the more links a page has, the less authority each one passes, because the same pool is split more ways. That is why a link from a focused, sparingly-linked page can be worth more than one buried among hundreds of links in a sprawling footer.

The nofollow chapter

In 2005, the major search engines introduced the nofollow link attribute to fight comment spam. A link marked rel="nofollow" told search engines, in effect, "don't pass authority through this link." That launched years of "PageRank sculpting," in which webmasters used nofollow to steer their internal authority toward the pages they cared about most. Google later changed how it handled nofollow — treating it as a hint rather than a strict command and adding the more specific sponsored and ugc values — but the underlying idea of consciously managing where your links send authority remains good practice.

Internal linking today

You no longer get a public number, but the crawl's core lesson is more relevant than ever: your internal links shape how authority and crawlers move through your own site. A few durable habits follow directly from it:

  • Link to your most important pages from your homepage and main navigation, so authority concentrates where it counts.
  • Keep key pages a few clicks from the homepage, not buried deep, so both users and crawlers reach them easily.
  • Use descriptive anchor text that tells search engines what the linked page is about.
  • Link out to genuinely helpful, trustworthy sources — good outbound links are a mark of a quality page, not a leak to be feared.

Crawl budget and modern crawlers

The crawl also touched on something search engines still care about deeply: how efficiently a site can be explored. Every search engine allocates a rough "crawl budget" to each site — how many pages its bots will fetch in a given period. A tangled structure, endless low-value URLs, or slow responses waste that budget and can leave important pages under-crawled. The old tool's habit of mapping every link on a page foreshadowed today's crawl-analysis tools, which flag redirect chains, orphaned pages, and crawl traps so that a site's authority and its crawlers both reach the pages that matter most.

From crawl to modern audits

The PageRank crawl was an early ancestor of today's site-audit and link-analysis tools, which map a site's entire internal and external link structure and flag where authority is trapped or wasted. If you want to turn this into action, our guide on checking a website's authority covers the modern equivalents, and what replaced PageRank explains the metrics they report.