A link, or hyperlink, is a reference in a web page that points to another resource, written in HTML as an <a> element with an href attribute [5]. In search engine optimisation, internal links connect pages on the same site and anchor text is the visible, clickable label. Search engines use links both to discover pages and to understand how relevant they are [1][2].
- Google finds most new pages by following links from pages it already knows [2].
- Only <a> elements with an href to a resolvable URL are reliably crawled [1].
- Anchor text should describe the target page; Google advises against generic text such as “click here” [1].
- Since 2020 Google treats nofollow, sponsored and ugc as hints rather than strict rules [7].
1Overview
Google discovers most new pages by following links from pages it already knows, and links are among the ways it judges what a page is about and how important it is [1][2].
Internal links are the part of this a site owner fully controls. They determine which pages crawlers can reach, how many paths lead to each page, and the words used to describe it.
2History
In 1998 Google’s founders described PageRank, a method of ranking pages using the link structure of the web, in which a link from one page to another counts as a kind of vote [3]. Google still lists link analysis and PageRank among its ranking systems, alongside many others [4].
In January 2005 Google introduced the rel="nofollow" attribute so that sites could mark links, particularly in blog comments, that they didn’t vouch for [6]. In September 2019 it added rel="sponsored" and rel="ugc", and announced that all three would be treated as hints; from 1 March 2020 that applied to crawling and indexing as well as ranking [7].
3Anatomy of a link
4Crawlable links
Google can reliably follow a link only if it is an <a> element with an href attribute containing a resolvable URL, absolute or relative [1].
| Markup | Crawlable? |
|---|---|
<a href="https://example.com/shoes"> | Yes [1] |
<a href="/shoes"> | Yes, relative URLs resolve [1] |
<a onclick="goTo('shoes')"> | No: no href [1] |
<span href="/shoes"> | No, not an <a> element [1] |
<a href="javascript:goTo()"> | No, not a URL [1] |
For single-page applications, Google recommends real URLs with the History API rather than fragment identifiers to load different content [15].
5Anchor text
Google recommends anchor text that is descriptive, reasonably concise and relevant to the page it points to and the page it is on. It should make sense out of context, and avoid generic phrases such as “click here” or “read more” [1].
Google also warns against stuffing anchors with keywords or making every link on a page use the same text; anchors should read naturally [1].
For image links, the image’s alt attribute acts as the anchor text [1].
6Internal linking and site structure
Google’s starter guide recommends linking to relevant pages within the site so readers, and crawlers, can find related content [16].
A page with no internal links pointing to it is often called an orphan page. Google can still find it through a sitemap, but sitemaps are recommended especially when pages are not well linked, which underlines that links are the primary path [10].
Breadcrumb navigation gives every page a link back up its hierarchy, and can be marked up with breadcrumb structured data so Google shows the trail in results [11].
7Broken links
A broken link points to a URL that returns an error, usually 404 Not Found. Google has said that 404s for pages that genuinely don’t exist don’t hurt the rest of a site’s rankings [12].
The cost is on the linking page: the visitor reaches a dead end and the link passes nothing. Pages that return 404 or 410 are dropped from the index over time, and pages that return 200 for missing content are treated as soft 404s [13].
Links that point to a redirect still work, Googlebot follows up to 10 hops, but each hop is an extra request, so internal links should point at the final URL [13][14].
8Link attributes
| Attribute | Use for |
|---|---|
| rel="sponsored" | Paid, affiliate or otherwise compensated links [8] |
| rel="ugc" | Links in user-generated content such as comments and forums [8] |
| rel="nofollow" | Other links you don’t want to vouch for [8] |
Values can be combined, for example rel="ugc nofollow"[8]. Using nofollow on internal links to “sculpt” authority is not something Google recommends; the attribute exists to describe a relationship [7].
9Link spam
Google’s spam policies prohibit link schemes intended to manipulate rankings, including buying or selling links that pass ranking credit, excessive link exchanges, and automated link creation [9].
10Common misconceptions
| Belief | What the sources say |
|---|---|
| 404s hurt the whole site | They don’t affect other pages’ rankings [12] |
| nofollow is a strict command | It is a hint for ranking, crawling and indexing [7] |
| Any clickable element is a link | Only <a href> is reliably followed [1] |
| A sitemap replaces internal links | Sitemaps help discovery, but links remain the main path [10][2] |
See also
- Links guide: diagrams and quick fixes
- Web crawling: how links feed discovery
- HTTP redirects: what happens when a link target moves
- Broken Link Checker: check one page’s links instantly
References
- [1]“Link best practices for Google”. Google Search Central. developers.google.com/search/docs/crawling-indexing/links-crawlable
- [2]“In-depth guide to how Google Search works”. Google Search Central. developers.google.com/search/docs/fundamentals/how-search-works
- [3]“The Anatomy of a Large-Scale Hypertextual Web Search Engine”. Brin & Page, Stanford, 1998. http://infolab.stanford.edu/~backrub/google.html
- [4]“A guide to Google Search ranking systems”. Google Search Central. developers.google.com/search/docs/appearance/ranking-systems-guide
- [5]“HTML Standard: Links”. WHATWG. html.spec.whatwg.org/multipage/links.html
- [6]“Preventing comment spam”. Official Google Blog, January 2005. googleblog.blogspot.com/2005/01/preventing-comment-spam.html
- [7]“Evolving “nofollow” – new ways to identify the nature of links”. Google Search Central Blog, September 2019. developers.google.com/search/blog/2019/09/evolving-nofollow-new-ways-to-identify
- [8]“Qualify your outbound links to Google”. Google Search Central. developers.google.com/search/docs/crawling-indexing/qualify-outbound-links
- [9]“Spam policies for Google web search”. Google Search Central. developers.google.com/search/docs/essentials/spam-policies
- [10]“Learn about sitemaps”. Google Search Central. developers.google.com/search/docs/crawling-indexing/sitemaps/overview
- [11]“Breadcrumb structured data”. Google Search Central. developers.google.com/search/docs/appearance/structured-data/breadcrumb
- [12]“Do 404s hurt my site?”. Google Search Central Blog, May 2011. developers.google.com/search/blog/2011/05/do-404s-hurt-my-site
- [13]“How HTTP status codes and network errors affect Google Search”. Google Search Central. developers.google.com/search/docs/crawling-indexing/http-network-errors
- [14]“Redirects and Google Search”. Google Search Central. developers.google.com/search/docs/crawling-indexing/301-redirects
- [15]“Understand the JavaScript SEO basics”. Google Search Central. developers.google.com/search/docs/crawling-indexing/javascript/javascript-seo-basics
- [16]“SEO Starter Guide”. Google Search Central. developers.google.com/search/docs/fundamentals/seo-starter-guide